haproxy

mirror of https://git.haproxy.org/git/haproxy.git/ synced 2025-08-06 23:27:04 +02:00

Author	SHA1	Message	Date
Christopher Faulet	533121a56e	MINOR: cache: Add global option to enable/disable zero-copy forwarding tune.cache.zero-copy-forwarding parameter can now be used to enable or disable the zero-copy fast-forwarding for the cache applet only. It is enabled ('on') by default. It can be disabled by setting the parameter to 'off'.	2023-12-06 10:24:41 +01:00
Christopher Faulet	ebead3c0a1	MEDIUM: cache: Add support for endp-to-endp fast-forwarding It is now possible to directly forward data to the opposite side from the cache applet. To do so, dedicated functions were added to fast-forward the payload part of the cached objects. Of course headers and trailers are still sent via the channel's buffer, using the HTX. When an object is delivered from the cache, once the applet reaches the HTX_CACHE_DATA state, it declares it can fast-forward data. From this point, all data are directly transferred to the oppposite side.	2023-12-06 10:24:41 +01:00
Christopher Faulet	5baa9ea168	MEDIUM: cache: Save body size of cached objects and track it on delivery We now save the body size of cached objets in the cache entry strucutre. In addition, the cache applet tracks the body part already sent. This will be mandatory to add support of endpoint-to-endpoint fast-forwarding in the cache applet.	2023-12-06 10:24:41 +01:00
Remi Tricot-Le Breton	23c810d042	BUG/MINOR: cache: Remove incomplete entries from the cache when stream is closed When a stream is interrupted by the client before the full answer is stored in the cache, we end up with an incomplete entry in the cache that cannot be overwritten until it "naturally" expires. In such a case, we call the cache filter's cache_store_strm_deinit callback without ever calling cache_store_http_end which means that the 'complete' flag is never set on the concerned cache_entry. This patch adds a check on the 'complete' flag in the strm_deinit callback and removes the entry from the cache if it is incomplete. A way to exhibit this bug is to try to get the same "big" response on multiple clients at the same time thanks to h2load for instance, and to interrupt the client side before the answer can be fully stored in the cache. This patch can be backported up to 2.4 but it will need some rework starting with branch 2.8 because of the latest cache changes.	2023-11-28 17:18:48 +01:00
Ilya Shipitsin	80813cdd2a	CLEANUP: assorted typo fixes in the code and comments This is 37th iteration of typo fixes	2023-11-23 16:23:14 +01:00
Willy Tarreau	3e913909e7	BUILD: cache: fix build error on older compilers pre-c99 compilers will fail to build the cache since commit `48f81ec09` ("MAJOR: cache: Delay cache entry delete in reserve_hot function") due to an int declaration in the for loop. No backport is needed.	2023-11-20 11:43:52 +01:00
Remi Tricot-Le Breton	f1f8e2b3df	DOC: cache: Specify when function expects a cache lock Some functions are built on the fact that the cache lock must be already taken by the caller. This patch adds this information in the functions' descriptions.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	45a2ff0f4a	MINOR: shctx: Remove 'use_shared_mem' variable This global variable was used to avoid using locks on shared_contexts in the unlikely case of nbthread==1. Since the locks do not do anything when USE_THREAD is not defined, it will be more beneficial to simply remove this variable and the systematic test on its value in the shared context locking functions.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	4fe6c1365d	MINOR: shctx: Remove redundant arg from free_block callback The free_block callback does not get called on blocks that are not row heads anymore so we don't need too shared_block parameters.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	48f81ec09d	MAJOR: cache: Delay cache entry delete in reserve_hot function A reference counter on the cache_entry was added in a previous commit. Its value is atomically increased and decreased via the retain_entry and release_entry functions. This is needed because of the latest cache and shared_context modifications that introduced two separate locks instead of the preexisting single shctx_lock one. With the new logic, we have two main blocks competing for the two locks: - the one in the http_action_req_cache_use that performs a lookup in the cache tree (locked by the cache lock) and then tries to remove the corresponding blocks from the shared_context's 'avail' list until the response is sent to the client by the cache applet, - the shctx_row_reserve_hot that traverses the 'avail' list and gives them back to the caller, while removing previous row heads from the cache tree Those two blocks require the two locks but one of them would take the cache lock first, and the other one the shctx_lock first, which would end in a deadlock without the current patch. The way this conflict is resolved in this patch is by ensuring that at least one of those uses works without taking the two locks at the same time. The solution found was to keep taking the two locks in the cache_use case. We first lock the cache to lookup for an entry and we then take the shctx lock as well to detach the corresponding blocks from the 'avail' list. The subtlety is that between the cache lookup and the actual locking of the shctx, another thread might have called the reserve_hot function in which we only take the shctx lock. In this function we traverse the 'avail' list to remove blocks that are then given to the caller. If one of those blocks corresponds to a previous row head, we call the 'free_blocks' callback that used to delete the cache entry from the tree. We now avoid deleting directly the cache entries in reserve_hot and we rather set the cache entries 'complete' param to 0 so that no other thread tries to work with this entry. This way, when we release the shctx lock in reserve_hot, the first thread that had performed the cache lookup and had found an entry that we just gave to another thread will see that the 'complete' field is 0 and it won't try to work with this response. The actual removal of entries from the cache tree will now be performed in the new 'reserve_finish' callback called at the end of the shctx_row_reserve_hot function. It will iterate on all the row head that were inserted in a dedicated list in the 'free_block' callback and perform the actual delete.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	11df806c88	MEDIUM: shctx: Descend shctx_lock calls into the shctx_row_reserve_hot Descend the shctx_lock calls into the shctx_row_reserve_hot so that the cases when we don't need to lock anything (enough space in the current row or not enough space in the 'avail' list) do not take the lock at all. In sh_ssl_sess_new_cb the lock had to be descended into sh_ssl_sess_store in order not to cover the shctx_row_reserve_hot call anymore.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	a29b073f26	MEDIUM: cache: Add refcount on cache_entry Add a reference counter on the cache_entry. Its value will be atomically increased and decreased via the retain_entry and release_entry functions. The release_entry function has two distinct versions, release_entry_locked and release_entry_unlocked that should be called when the cache lock is already taken in write mode or not (respectively). In the unlocked case the cache lock will only be taken in write mode on the last reference of the entry (before calling delete_entry). This allows to limit the amount of times when we need to take the cache lock during a release operation.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	ed35b9411a	MEDIUM: cache: Switch shctx spinlock to rwlock and restrict its scope Since a lock on the cache tree was added in the latest cache changes, we do not need to use the shared_context's lock to lock more than pure shared_context related data anymore. This already existing lock will now only cover the 'avail' list from the shared_context. It can then be changed to a rwlock instead of a spinlock because we might want to only run through the avail list sometimes. Apart form changing the type of the shctx lock, the main modification introduced by this patch is to limit the amount of code covered by the shctx lock. This lock does not need to cover any code strictly related to the cache tree anymore.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	a0d7c290ec	MINOR: cache: Use dedicated trash for "show cache" cli command After the latest changes in the cache/shared_context mechanism, the cache and shared_context logic were decorrelated and in some unlikely cases we might end up using the "show cache" command while some regular cache processing is occurring (a response being stored in the cache for instance). In such a case, because we used the same 'trash' buffer in those two contexts, we could end up with the contents of a response in the ouput of the "show cache" command. This patch fixes this problem by allocating a dedicated trash for the CLI command.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	3831d8454f	MEDIUM: shctx: Remove 'hot' list from shared_context The "hot" list stored in a shared_context was used to keep a reference to shared blocks that were currently being used and were thus removed from the available list (so that they don't get reused for another cache response). This 'hot' list does not ever need to be shared across threads since every one of them only works on their current row. The main need behind this 'hot' list was to detach the corresponding blocks from the 'avail' list and to have a known list root when calling list_for_each_entry_from in shctx_row_data_append (for instance). Since we actually never need to iterate over all members of the 'hot' list, we can remove it and replace the inc_hot/dec_hot logic by a detach/reattach one.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	bd24118212	MEDIUM: cache: Use rdlock on cache in cache_use When looking for a valid entry in the cache tree in http_action_req_cache_use, we do not need to delete an expired entry at once because even if an expired entry exists, since the request will be forwarded to the server, then the expired entry will be overwritten when the updated response is seen. We can then use a simpler rdlock during cache_use operation.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	0dfb57bbf9	MINOR: cache: Add option to avoid removing expired entries in lookup function Any lookup in the cache tree done through entry_exist or secondary_entry_exist functions could end up deleting the corresponding entry if it is expired which prevents from using a rdlock on code paths that would just perform a lookup on the tree (in http_action_req_cache_use for instance). Adding a 'delete_expired' boolean as a parameter allows for "pure" lookups and thus it will allow to perform operations on the tree that simply require a rdlock instead of a "heavier" wrlock.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	ff3cb6dad4	MINOR: cache: Remove expired entry delete in "show cache" command The "show cache" CLI command iterates over all the entries of the cache tree and it used this opportunity to remove expired entries from the cache. This behavior was completely undocumented and does not seem that necessary. By removing it we can take the cache lock in read mode only which limits the impact on the other threads.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	ac9c49b40d	MEDIUM: cache: Use dedicated cache tree lock alongside shctx lock Every use of the cache tree was covered by the shctx lock even when no operations were performed on the shared_context lists (avail and hot). This patch adds a dedicated RW lock for the cache so that blocks of code that work on the cache tree only can use this lock instead of the superseding shctx one. This is useful for operations during which the concerned blocks are already in the hot list. When the two locks need to be taken at the same time, in http_action_req_cache_use and in shctx_row_reserve_hot, the shctx one must be taken first. A new parameter needed to be added to the shared_context's free_block callback prototype so that cache_free_block can take the cache lock and release it afterwards.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	81d8014af8	MINOR: shctx: Remove explicit 'from' param from shctx_row_data_append This parameter is not necessary since the first element of a row always has a pointer to the row's tail.	2023-11-16 19:35:10 +01:00
Remi Tricot-Le Breton	a5e96425a2	MEDIUM: cache: Add "Origin" header to secondary cache key This patch add a hash of the Origin header to the cache's secondary key. This enables to manage store responses that have a "Vary: Origin" header in the cache when vary is enabled. This cannot be considered as a means to manage CORS requests though, it only processes the Origin header and hashes the presented value without any form of URI normalization. This need was expressed by Philipp Hossner in GitHub issue #251. Co-Authored-by: Philipp Hossner <philipp.hossner@posteo.de>	2023-10-05 10:53:54 +02:00
Remi Tricot-Le Breton	e03d060aa3	MINOR: cache: Change hash function in default normalizer used in case of "vary" When building the secondary signature for cache entries when vary is enabled, the referer part of the signature was a simple crc32 of the first referer header. This patch changes it to a 64bits hash based of xxhash algorithm with a random seed built during init. This will prevent "malicious" hash collisions between entries of the cache.	2023-09-06 16:11:31 +02:00
Christopher Faulet	d6f0557deb	BUG/MEDIUM: cache: Don't request more room than the max allowed Since a recent change on the SC API, a producer must specify the amount of free space it needs to progress when it is blocked. But, it must take care to never exceed the maximum size allowed in the buffer. Otherwise, the stream is freezed because it cannot reach the condition to unblock the producer. In this context, there is a bug in the cache applet when it fails to dump a message. It may request more space than allowed. It happens when the cached object is too big. It is a 2.8-specific bug. No backport needed.	2023-05-09 11:53:28 +02:00
Christopher Faulet	7b3d38a633	MEDIUM: tree-wide: Change sc API to specify required free space to progress sc_need_room() now takes the required free space to receive more data as parameter. All calls to this function are updated accordingly. For now, this value is set but not used. When we are waiting for a buffer, 0 is used. So we expect to be unblocked ASAP. However this must be reviewed because SC_FL_NEED_BUF is probably enough in this case and this flag is already set if the input buffer allocation fails.	2023-05-05 15:44:23 +02:00
Ilya Shipitsin	2ca01589a0	CLEANUP: use "offsetof" where appropriate let's use the C library macro "offsetof"	2023-04-16 09:58:49 +02:00
Christopher Faulet	f8130b2de2	MEDIUM: cache: Use the sedesc to report and detect end of processing We now try, as far as possible, to rely on the SE descriptor to detect end of processing. Idea is to no longer rely on the channel or the SC to do so. First, we now set SE_FL_EOS instead of calling and cf_shutr() to report the end of the stream. It happens when the response is fully sent (SE_FL_EOI is already set in this case) or when an error is reported. In this last case, SE_FL_ERROR is also set. Thanks to this change, it is now possible to detect the applet must only consume the request waiting for the upper layer releases it. So, if SE_FL_EOS or SE_FL_ERROR are set, it means the reponse was fully handled. And if SE_FL_SHR or SE_FL_SHW are set, it means the applet was released by upper layer and is waiting to be freed.	2023-04-05 08:57:05 +02:00
Christopher Faulet	92297749e1	MINOR: applet: No longer set EOI on the SC Thanks to the previous patch, it is now possible for applets to not set the CF_EOI flag on the channels. On this point, the applets get closer to the muxes.	2023-04-05 08:57:05 +02:00
Christopher Faulet	b08c5259eb	MINOR: stconn: Always report READ/WRITE event on shutr/shutw It was done by hand by callers when a shutdown for read or write was performed. It is now always handled by the functions performing the shutdown. This way the callers don't take care of it. This will avoid some bugs.	2023-02-22 15:59:16 +01:00
Remi Tricot-Le Breton	25917cdb12	BUG/MINOR: cache: Check cache entry is complete in case of Vary Before looking for a secondary cache entry for a given request we checked that the first entry was complete, which might prevent us from using a valid entry if the first one with the same primary key is not full yet. Likewise, if the primary entry is complete but not the secondary entry we try to use, we might end up using a partial entry from the cache as a response. This bug was raised in GitHub #2048. It can be backported up to branch 2.4.	2023-02-21 18:35:48 +01:00
Remi Tricot-Le Breton	879debeecb	BUG/MINOR: cache: Cache response even if request has "no-cache" directive Since commit `cc9bf2e5f` "MEDIUM: cache: Change caching conditions" responses that do not have an explicit expiration time are not cached anymore. But this mechanism wrongly used the TX_CACHE_IGNORE flag instead of the TX_CACHEABLE one. The effect this had is that a cacheable response that corresponded to a request having a "Cache-Control: no-cache" for instance would not be cached. Contrary to what was said in the other commit message, the "checkcache" option should not be impacted by the use of the TX_CACHEABLE flag instead of the TX_CACHE_IGNORE one. The response is indeed considered as not cacheable if it has no expiration time, regardless of the presence of a cookie in the response. This should fix GitHub issue #2048. This patch can be backported up to branch 2.4.	2023-02-21 18:35:41 +01:00
Willy Tarreau	9b5d57dfd5	BUG/MEDIUM: cache: use the correct time reference when comparing dates The cache makes use of dates advertised by external components, such as "last-modified" or "date". As such these are wall-clock dates, and not internal dates. However, all comparisons are mistakenly made based on the internal monotonic date which is designed to drift from the wall clock one in order to catch up with stolen time (which can sometimes be intense in VMs). As such after some run time some objects may fail to validate or fail to expire depending on the direction of the drift. This is particularly visible when applying an offset to the internal time to force it to wrap soon after startup, as it will be shifted up to 49.7 days in the future depending on the current date; what happens in this case is that the reg-test "cache_expires.vtc" fails on the 3rd test by returning stale contents from the cache at the date of this commit. It is really important that all external dates are compared against "date" and not "now" for this reason. This fix needs to be backported to all versions.	2023-02-08 11:10:33 +01:00
Christopher Faulet	6e1bbc446b	REORG: channel: Rename CF_READ_NULL to CF_READ_EVENT CF_READ_NULL flag is not really useful and used. It is a transient event used to wakeup the stream. As we will see, all read events on a channel may be resumed to only one and are all used to wake up the stream. In this patch, we introduce CF_READ_EVENT flag as a replacement to CF_READ_NULL. There is no breaking change for now, it is just a rename. Gradually, other read events will be merged with this one.	2023-01-09 18:41:08 +01:00
Willy Tarreau	c12b321661	CLEANUP: applet: rename appctx_cs() to appctx_sc() It returns a stream connector, not a conn_stream anymore, so let's fix its name.	2022-05-27 19:33:35 +02:00
Willy Tarreau	c5ddd9fd80	CLEANUP: cache: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	cb086c6de1	REORG: stconn: rename conn_stream.{c,h} to stconn.{c,h} There's no more reason for keepin the code and definitions in conn_stream, let's move all that to stconn. The alphabetical ordering of include files was adjusted.	2022-05-27 19:33:35 +02:00
Willy Tarreau	5edca2f0e1	REORG: rename cs_utils.h to sc_strm.h This file contains all the stream-connector functions that are specific to application layers of type stream. So let's name it accordingly so that it's easier to figure what's located there. The alphabetical ordering of include files was preserved.	2022-05-27 19:33:35 +02:00
Willy Tarreau	f61dd19284	CLEANUP: stconn: rename cs_{shut,chk}* to sc_* This applies the following renaming: cs_shutr() -> sc_shutr() cs_shutw() -> sc_shutw() cs_chk_rcv() -> sc_chk_rcv() cs_chk_snd() -> sc_chk_snd() cs_must_kill_conn() -> sc_must_kill_conn()	2022-05-27 19:33:35 +02:00
Willy Tarreau	a0b58b537d	CLEANUP: stconn: rename cs_{new,create,free,destroy}_* to sc_* This renames the following functions: cs_new_from_endp() -> sc_new_from_endp() cs_new_from_strm() -> sc_new_from_strm() cs_new_from_check() -> sc_new_from_check() cs_applet_create() -> sc_applet_create() cs_destroy() -> sc_destroy() cs_free() -> sc_free()	2022-05-27 19:33:35 +02:00
Willy Tarreau	99615ed85d	CLEANUP: stconn: rename cs_rx_room_{blk,rdy} to sc_{need,have}_room() The new name mor eclearly indicates that a stream connector cannot make any more progress because it needs room in the channel buffer, or that it may be unblocked because the buffer now has more room available. The testing function is sc_waiting_room(). This is mostly used by applets. Note that the flags will change soon.	2022-05-27 19:33:35 +02:00
Willy Tarreau	8e7c6e6907	CLEANUP: stconn: rename cs_appctx() to sc_appctx() Nothing special, just s/cs/sc/, roughly 50-60 entries.	2022-05-27 19:33:34 +02:00
Willy Tarreau	ea27f48c5a	CLEANUP: stconn: rename cs_{check,strm,strm_task} to sc_strm_* These functions return the app-layer associated with an stconn, which is a check, a stream or a stream's task. They're used a lot to access channels, flags and for waking up tasks. Let's just name them appropriately for the stream connector.	2022-05-27 19:33:34 +02:00
Willy Tarreau	40a9c32e3a	CLEANUP: stconn: rename cs_{i,o}{b,c} to sc_{i,o}{b,c} We're starting to propagate the stream connector's new name through the API. Most call places of these functions that retrieve the channel or its buffer are in applets. The local variable names are not changed in order to keep the changes small and reviewable. There were ~92 uses of cs_ic(), ~96 of cs_oc() (due to co_get() being less factorizable than ci_put), and ~5 accesses to the buffer itself.	2022-05-27 19:33:34 +02:00
Willy Tarreau	d0a06d52f4	CLEANUP: applet: use applet_put() everywhere possible This applies the change so that the applet code stops using ci_putchk() and friends everywhere possible, for the much saferapplet_put() instead. The change is mechanical but large. Two or three functions used to have no appctx and a cs derived from the appctx instead, which was a reminiscence of old times' stream_interface. These were simply changed to directly take the appctx. No sensitive change was performed, and the old (more complex) API is still usable when needed (e.g. the channel is already known). The change touched roughly a hundred of locations, with no less than 124 lines removed. It's worth noting that the stats applet, the oldest of the series, could get a serious lifting, as it's still very channel-centric instead of propagating the appctx along the chain. Given that this code doesn't change often, there's no emergency to clean it up but it would look better.	2022-05-27 19:33:34 +02:00
Willy Tarreau	026e8fb290	CLEANUP: stconn: tree-wide rename stconn states CS_ST/SB_* to SC_ST/SB_* This also follows the natural naming. There are roughly 238 changes, all totally trivial. conn_stream-t.h has become completely void of any "conn_stream" related stuff now (except its name).	2022-05-27 19:33:34 +02:00
Willy Tarreau	7cb9e6c6ba	CLEANUP: stream: rename "csf" and "csb" to "scf" and "scb" These are the stream connectors, let's give them consistent names. The patch is large (405 locations) but totally trivial.	2022-05-27 19:33:34 +02:00
Willy Tarreau	4596fe20d9	CLEANUP: conn_stream: tree-wide rename to stconn (stream connector) This renames the "struct conn_stream" to "struct stconn" and updates the descriptions in all comments (and the rare help descriptions) to "stream connector" or "connector". This touches a lot of files but the change is minimal. The local variables were not even renamed, so there's still a lot of "cs" everywhere.	2022-05-27 19:33:34 +02:00
Willy Tarreau	d869e13ed8	CLEANUP: applet: rename the sedesc pointer from "endp" to "sedesc" Now at least it makes it obvious that it's the stream endpoint descriptor and not an endpoint. There were few changes thanks to the previous refactor of the flags.	2022-05-27 19:33:34 +02:00
Willy Tarreau	b605c4213f	CLEANUP: conn_stream: rename the stream endpoint flags CS_EP_* to SE_FL_* Let's now use the new flag names for the stream endpoint.	2022-05-27 19:33:34 +02:00
Willy Tarreau	d56377c5eb	CLEANUP: conn_stream: apply endp_flags.cocci tree-wide This changes all main uses of endp->flags to the se_fl_() equivalent by applying coccinelle script endp_flags.cocci. The se_fl_() functions themselves were manually excluded from the change, of course. Note: 144 locations were touched, manually reviewed and found to be OK. The script was applied with all includes: spatch --in-place --recursive-includes -I include --sp-file $script $files	2022-05-27 19:33:34 +02:00
Willy Tarreau	0698c80a58	CLEANUP: applet: remove the unneeded appctx->owner This one is the pointer to the conn_stream which is always in the endpoint that is always present in the appctx, thus it's not needed. This patch removes it and replaces it with appctx_cs() instead. A few occurences that were using __cs_strm(appctx->owner) were moved directly to appctx_strm() which does the equivalent.	2022-05-13 14:28:48 +02:00

1 2 3 4 5 ...

275 Commits