haproxy

mirror of https://git.haproxy.org/git/haproxy.git/ synced 2025-08-12 10:06:58 +02:00

Author	SHA1	Message	Date
Frédéric Lécaille	a2d8ad20a3	MINOR: quic: Replace MT_LISTs by LISTs for RX packets. Replace ->rx.pqpkts quic_enc_level struct member MT_LIST by an LIST. Same thing for ->list quic_rx_packet struct member MT_LIST. Update the code consequently. This was a reminisence of the multithreading support (several threads by connection). Must be backported to 2.6	2022-08-23 17:55:02 +02:00
Frédéric Lécaille	b8047de11a	BUG/MINOR: quic: Safer QUIC frame builders Do not rely on the fact the callers of qc_build_frm() handle their buffer passed to function the correct way (without leaving garbage). Make qc_build_frm() update the buffer passed as argument only if the frame it builds is well formed. As far as I sse, there is no such callers which does not handle carefully such buffers. Must be backported to 2.6.	2022-08-23 17:40:09 +02:00
Frédéric Lécaille	a8a6043240	BUG/MINOR: quic: Wrong list_for_each_entry() use when building packets from qc_do_build_pkt() This is list_for_each_entry_safe() which must be used if we want to delete elements inside its code block. This could explain that some frames which were not built were added to packets with a NULL ->pkt member. Thank you to Tristan for having reported this issue through backtraces in GH #1808 Must be backported to 2.6.	2022-08-23 12:06:40 +02:00
Frédéric Lécaille	da9c441886	BUG/MINOR: quix: Memleak for non in flight TX packets First, these packets must not be inserted in the tree of TX packets. They are never explicitely acknowledged (for instance an ACK only packet will never be acknowledged). Furthermore, if taken into an account these packets may uselessly disturb the congestion control. We do not care if they are lost or not. Furthermore as the ->in_fligh_len member value is null they were not released by qc_release_lost_pkts() which rely on these values to decide to release the allocated memory for such packets. Must be backported to 2.6.	2022-08-22 19:06:08 +02:00
Emeric Brun	8032a276ce	BUG/MAJOR: mworker: fix infinite loop on master with no proxies. The master is re-exec with an empty proxies list if no master CLI is configured. This results in infinite loop since last patch: `3b68b602` ("BUG/MAJOR: log-forward: Fix log-forward proxies not fully initialized") This patch avoid to loop again on log-forward proxies list if empty. This patch should be backported until v2.3	2022-08-22 13:09:29 +02:00
Willy Tarreau	f1cfd9bc97	MINOR: cpu-map: remove obsolete diag warning about combined ranges We used to emit a diag warning in case ranges were used both with the process and thread part of a thread spec. Now with groups it's not longer a problem, so let's just kill this warning.	2022-08-22 10:46:13 +02:00
Willy Tarreau	3cd71acd06	BUG/MEDIUM: cpu-map: fix thread 1's affinity affecting all threads Since 2.7-dev2 with commit `5b09341c02` ("MEDIUM: cpu-map: replace the process number with the thread group number"), the thread group has replaced the process number in the "cpu-map" directive. In part due to a design limit in 2.4 and 2.5, a special case was made of thread 1 in commit `bda7c1decd` ("MEDIUM: config: simplify cpu-map handling"), because there was no other location to store a single-threaded setup's mask by then. The combination of the two resulted in a problem with thread groups, by which as soon as one line exhibiting thread number 1 alone was found in a config, the mask would be applied to all threads in the group. The loop was reworked to avoid this obsolete special case, and was factored for better legibility. One obsolete comment about nbproc was also removed. No backport is needed.	2022-08-22 10:38:00 +02:00
Frédéric Lécaille	ea4a5cbbdf	BUG/MINOR: mux-quic: Fix memleak on QUIC stream buffer for unacknowledged data Some clients send CONNECTION_CLOSE frame without acknowledging the STREAM data haproxy has sent. In this case, when closing the connection if there were remaining data in QUIC stream buffers, they were not released. Add a <closing> boolean option to qc_stream_desc_free() to force the stream buffer memory releasing upon closing connection. Thank you to Tristan for having reported such a memory leak issue in GH #1801. Must be backported to 2.6.	2022-08-20 19:08:31 +02:00
William Lallemand	62c0b99e3b	MINOR: ssl/cli: implement "add ssl ca-file" In ticket #1805 an user is impacted by the limitation of size of the CLI buffer when updating a ca-file. This patch allows a user to append new certificates to a ca-file instead of trying to put them all with "set ssl ca-file" The implementation use a new function ssl_store_dup_cafile_entry() which duplicates a cafile_entry and its X509_STORE. ssl_store_load_ca_from_buf() was modified to take an apped parameter so we could share the function for "set" and "add".	2022-08-19 19:58:53 +02:00
William Lallemand	d4774d3cfa	MINOR: ssl: handle ca-file appending in cafile_entry In order to be able to append new CA in a cafile_entry, ssl_store_load_ca_from_buf() was reworked and a "append" parameter was added. The function is able to keep the previous X509_STORE which was already present in the cafile_entry.	2022-08-19 19:58:53 +02:00
William Lallemand	ec7eb59d20	BUG/MINOR: ssl/cli: error when the ca-file is empty "set ssl ca-file" does not return any error when a ca-file is empty or only contains comments. This could be a problem is the file was malformated and did not contain any PEM header. It must be backported as far as 2.5.	2022-08-19 19:56:53 +02:00
Frédéric Lécaille	86a53c5669	MINOR: quic: Add reusable cipher contexts for header protection Implement quic_tls_rx_hp_ctx_init() and quic_tls_tx_hp_ctx_init() to initiliaze such header protection cipher contexts for each RX and TX parts and for each packet number spaces, only one time by connection. Make qc_new_isecs() call these two functions to initialize the cipher contexts of the Initial secrets. Same thing for ha_quic_set_encryption_secrets() to initialize the cipher contexts of the subsequent derived secrets (ORTT, 1RTT, Handshake). Modify qc_do_rm_hp() and quic_apply_header_protection() to reuse these cipher contexts. Note that there is no need to modify the key update for the header protection. The header protection secrets are never updated.	2022-08-19 18:31:59 +02:00
Emeric Brun	a8942cd9c4	BUG/MAJOR: log-forward: Fix ssl layer not initialized on bind even if configured Since commit `2071a99df` ("MINOR: listener/ssl: set the SSL xprt layer only once the whole config is known") the xprt is initialized for ssl directly from a generic funtion used to parse bind args. But the 'bind' lines from 'log-forward' sections were forgotten in commit `55f0f7bb5` ("MINOR: config: use the new bind_parse_args_list() to parse a "bind" line"). This patch re-works 'log-forward' section parsing to use the generic function to parse bind args and fix the issue. Since the generic way to parse was introduced in 2.6, this patch should be backported as far as this version.	2022-08-19 16:09:06 +02:00
Emeric Brun	3b68b60261	BUG/MAJOR: log-forward: Fix log-forward proxies not fully initialized Some initialisation for log forward proxies was missing such as ssl configuration on 'log-forward's 'bind' lines. After the loop on the proxy initialization code for proxies present in the main proxies list, this patch force to loop again on this code for proxies present in the log forward proxies list. Those two lists should be merged. This will be part of a global re-work of proxy initialization including peers proxies and resolver proxies. This patch was made in first attempt to fix the bug and to facilitate the backport on older branches waiting for a cleaner re-work on proxies initialization on the dev branch. This patch should be backported as far as 2.3.	2022-08-19 16:08:03 +02:00
Frédéric Lécaille	a846a17fde	MINOR: quic: Trace fix in qc_release_frm() This wrong trace came with this commit: "BUG/MINOR: quic: Possible crashes when dereferencing ->pkt quic_frame struct member" In qc_release_frm() we mark frames as acked. Nothing to see with references to frames. Thank you to Willy for having caught this one. Must be backported to 2.6 as these traces arrived with a bug fix to be backported to 2.6.	2022-08-19 12:15:05 +02:00
Frédéric Lécaille	e4c3074c00	MINOR: quic: Add the QUIC connection to mux traces This should help for debugging purpose. Should be backported to 2.6	2022-08-19 12:02:29 +02:00
Frédéric Lécaille	b827840b42	BUG/MINOR: quic: Wrong splitted duplicated frames handling When duplicated frames are splitted, we must propagate this information to the new allocated frame and add a reference to this new frame to the reference list of the original frame. Must be backported to 2.6	2022-08-19 10:10:43 +02:00
Frédéric Lécaille	2f16348d24	MINOR: quic: Add frame addresses to QUIC_EV_CONN_PRSAFRM event traces This should be useful to diagnose some issues. Should be backported to 2.6.	2022-08-19 09:59:07 +02:00
Frédéric Lécaille	1ba25c244e	BUG/MINOR: quic: Possible crashes when dereferencing ->pkt quic_frame struct member This was done at several places. First in qc_requeue_nacked_pkt_tx_frms. This aim of this function is, if needed, to requeue all the TX frames of a lost <pkt> packet passed as argument and detach them from this packet they have been sent from. They are possible cases where the frm->pkt quic_frame struct member could be NULL, as a result of a duplication of an original frame by qc_dup_pkt_frms(). This function adds the duplicated frame to the original frame reference list: LIST_APPEND(&origin->reflist, &dup_frm->ref); But, in this function, the packet which contains the frame is the one which is passed as argument (for debug purpose). So let us prefer using this variable. Also do not dereference this ->pkt quic_frame member in qc_release_frm() and qc_frm_unref() and add a trace to catch the frame with a null ->pkt member. They are logically frames which have not already been sent. Thank you to Tristan for having reported such crashes in GH #1808. Must be backported to 2.6	2022-08-19 09:58:28 +02:00
Willy Tarreau	473e0e54f5	BUG/MINOR: mux-h2: send a CANCEL instead of ES on truncated writes If a POST upload is cancelled after having advertised a content-length, or a response body is truncated after a content-length, we're not allowed to send ES because in this case the total body length must exactly match the advertised value. Till now that's what we were doing, and that was causing the other side (possibly haproxy) to respond with an RST_STREAM PROTOCOL_ERROR due to "ES on DATA frame before content-length". We can behave a bit cleaner here. Let's detect that we haven't sent everything, and send an RST_STREAM(CANCEL) instead, which is designed exactly for this purpose. This patch could be backported to older versions but only a little bit of exposure to make sure it doesn't wake up a bad behavior somewhere. It relies on the following previous commit: "MINOR: mux-h2: make streams know if they need to send more data"	2022-08-19 08:03:53 +02:00
Willy Tarreau	4877045f1d	MINOR: mux-h2: make streams know if they need to send more data H2 streams do not even know if they are expected to send more data or not, which is problematic when closing because we don't know if we're closing too early or not. Let's start by adding a new stream flag "H2_SF_MORE_HTX_DATA" to indicate this on the tx path.	2022-08-19 08:03:53 +02:00
Willy Tarreau	ed2b9d9f27	MINOR: mux-h2/traces: report transition to SETTINGS1 before not after Traces indicating "switching to XXX" generally apply before the transition so that the current connection state is visible in the trace. SETTINGS1 was incorrect in this regard, with the trace being emitted after. Let's fix this. No need to backport this, as this is purely cosmetic.	2022-08-19 08:03:53 +02:00
Willy Tarreau	0f45871344	BUG/MEDIUM: mux-h2: do not fiddle with ->dsi to indicate demux is idle When switching to H2_CS_FRAME_H, we do not want to present the previous frame's state, flags, length etc in traces, or we risk to confuse the analysis, making the reader think that the header information presented is related to the new frame header being analysed. A naive approach could have consisted in simply relying on the current parser state (FRAME_H being that state), but traces are emitted before switching the state, so traces cannot rely on this. This was initially addressed by commit `73db434f7` ("MINOR: h2/trace: report the frame type when known") which used to set dsi to -1 when the connection becomes idle again, but was accidentally broken by commit `5112a603d` ("BUG/MAJOR: mux_h2: Don't consume more payload than received for skipped frames") which moved dsi after calling the trace function. But in both cases there's problem with this approach. If an RST or WU frame cannot be uploaded due to a busy mux, and at the same time we complete processing on a perfect end of frame with no single new frame header, we can leave the demux loop with dsi=-1 and with RST or WU to be sent, and these ones will be sent for stream ID -1. This is what was reported in github issue #1830. This can be reproduced with a config chaining an h1->h2 proxy to an empty h2 frontend, and uploading a large body such as below: $ (printf "POST / HTTP/1.1\r\nContent-length: 1000000000\r\n\r\n"; cat /dev/zero) \| nc 0 4445 > /dev/null This shows that we must never affect ->dsi which must always remain valid, and instead we should set "something else". That something else could be served by the demux frame type, but that one also needs to be preserved for the RST_STREAM case. Instead, let's just add a connection flag to say that the demuxing is in progress. This will be set once a new demux header is set and reset after the end of a frame. This way the trace subsystem can know that dft/dfl must not be displayed, without affecting the logic relying on such values. Given that the commits above are old and were backported to 1.8, this new one also needs to be backported as far as 1.8. Many thanks to David le Blanc (@systemmonkey42) for spotting, reporting, capturing and analyzing this bug; his work permitted to quickly spot the problem.	2022-08-19 08:03:53 +02:00
Willy Tarreau	1addf8b777	BUG/MEDIUM: cli: always reset the service context between commands Erwan Le Goas reported that chaining certain commands on the CLI would systematically crash the process; for example, "show version; show sess". This happened since the conversion of cli context to appctx->svcctx, because if applet_reserve_svcctx() is called a first time for a tiny context, it's allocated in-situ, and later a keyword that wants a larger one will see that it's not null and will reuse it and will overwrite the end of the first one's context. What is missing is a reset of the svcctx when looping back to CLI_ST_GETREQ. This needs to be backported to 2.6, and relies on previous commit "MINOR: applet: add a function to reset the svcctx of an applet".	2022-08-18 18:16:36 +02:00
Willy Tarreau	1cc08a33e1	MINOR: applet: add a function to reset the svcctx of an applet The CLI needs to reset the svcctx between commands, and there was nothing done to handle this. Let's add appctx_reset_svcctx() to do that, it's the closing equivalent of appctx_reserve_svcctx(). This will have to be backported to 2.6 as it will be used by a subsequent patch to fix a bug.	2022-08-18 18:16:36 +02:00
Amaury Denoyelle	115ccce867	MEDIUM: h3: concatenate multiple cookie headers As specified by RFC 9114, multiple cookie headers must be concatenated into a single entry before passing it to a HTTP/1.1 connection. To implement this, reuse the same function as already used for HTTP/2 module. This should answer to feature requested in github issue #1818.	2022-08-18 16:13:33 +02:00
Amaury Denoyelle	2c5a7ee333	REORG: h2: extract cookies concat function in http_htx As specified by RFC 7540, multiple cookie headers are merged in a single entry before passing it to a HTTP/1.1 connection. This step is implemented during headers parsing in h2 module. Extract this code in the generic http_htx module. This will allow to reuse it quickly for HTTP/3 implementation which has the same requirement for cookie headers.	2022-08-18 16:13:33 +02:00
Amaury Denoyelle	704675656b	BUG/MEDIUM: quic: fix crash on MUX send notification MUX notification on TX has been edited recently : it will be notified only when sending its own data, and not for example on retransmission by the quic-conn layer. This is subject of the patch : `b29a1dc2f4` BUG/MINOR: quic: do not notify MUX on frame retransmit A new flag QUIC_FL_CONN_RETRANS_LOST_DATA has been introduced to differentiate qc_send_app_pkts invocation by MUX and directly by the quic-conn layer in quic_conn_app_io_cb(). However, this is a first problem as internal quic-conn layer usage is not limited to retransmission. For example for NEW_CONNECTION_ID emission. Another problem much important is that send functions are also called through quic_conn_io_cb() which has not been protected from MUX notification. This could probably result in crash when trying to notify the MUX. To fix both problems, quic-conn flagging has been inverted : when used by the MUX, quic-conn is flagged with QUIC_FL_CONN_TX_MUX_CONTEXT. To improve the API, MUX must now used qc_send_mux which ensure the flag is set. qc_send_app_pkts is now static and can only be used by the quic-conn layer. This must be backported wherever the previously mentionned patch is.	2022-08-18 11:33:22 +02:00
Frédéric Lécaille	4173a39c1f	BUG/MINOR: quic: Missing initializations for ducplicated frames. When duplication frames in qc_dup_pkt_frms(), ->pkt member was not correctly initialized (copied from the original frame). This could not have any impact because this member is initialized whe the frame is added to a packet. This was also the case for ->flags. Also replace the pool_zalloc() call by a call to pool_alloc(). Must be backported to 2.6.	2022-08-18 10:28:31 +02:00
Mateusz Malek	4b85a963be	BUG/MEDIUM: http-ana: fix crash or wrong header deletion by http-restrict-req-hdr-names When using `option http-restrict-req-hdr-names delete`, HAproxy may crash or delete wrong header after receiving request containing multiple forbidden characters in single header name; exact behavior depends on number of request headers, number of forbidden characters and position of header containing them. This patch fixes GitHub issue #1822. Must be backported as far as 2.2 (buggy feature got included in 2.2.25, 2.4.18 and 2.5.8).	2022-08-17 15:52:17 +02:00
Amaury Denoyelle	b29a1dc2f4	BUG/MINOR: quic: do not notify MUX on frame retransmit On STREAM emission, quic-conn notifies MUX through a callback named qcc_streams_sent_done(). This also happens on retransmission : in this case offset are examined and notification is ignored if already seen. However, this behavior has slightly changed since `e53b489826` BUG/MEDIUM: mux-quic: fix server chunked encoding response Indeed, if offset diff is NULL, frame is now not ignored. This is to support FIN notification with a final empty STREAM frame. A side-effect of this is that if the last stream frame is retransmitted, it won't be ignored in qcc_streams_sent_done(). In most cases, this side-effect is harmless as qcs instance will soon be freed after being closed. But if qcs is still alive, this will cause a BUG_ON crash as it is considered as locally closed. This bug depends on delay condition and seems to be extremely rare. But it might be the reason for a crash seen on interop with s2n client on http3 testcase : FATAL: bug condition "qcs->st == QC_SS_CLO" matched at src/mux_quic.c:372 call trace(16): \| 0x558228912b0d [b8 01 00 00 00 c6 00 00]: main-0x1c7878 \| 0x558228917a70 [48 8b 55 d8 48 8b 45 e0]: qcc_streams_sent_done+0xcf/0x355 \| 0x558228906ff1 [e9 29 05 00 00 48 8b 05]: main-0x1d3394 \| 0x558228907cd9 [48 83 c4 10 85 c0 0f 85]: main-0x1d26ac \| 0x5582289089c1 [48 83 c4 50 85 c0 75 12]: main-0x1d19c4 \| 0x5582288f8d2a [48 83 c4 40 48 89 45 a0]: main-0x1e165b \| 0x5582288fc4cc [89 45 b4 83 7d b4 ff 74]: qc_send_app_pkts+0xc6/0x1f0 \| 0x5582288fd311 [85 c0 74 12 eb 01 90 48]: main-0x1dd074 \| 0x558228b2e4c1 [48 c7 c0 d0 60 ff ff 64]: run_tasks_from_lists+0x4e6/0x98e \| 0x558228b2f13f [8b 55 80 29 c2 89 d0 89]: process_runnable_tasks+0x7d6/0x84c \| 0x558228ad9aa9 [8b 05 75 16 4b 00 83 f8]: run_poll_loop+0x80/0x48c \| 0x558228ada12f [48 8b 05 aa c5 20 00 48]: main-0x256 \| 0x7ff01ed2e609 [64 48 89 04 25 30 06 00]: libpthread:+0x8609 \| 0x7ff01e8ca163 [48 89 c7 b8 3c 00 00 00]: libc:clone+0x43/0x5e To reproduce it locally, code was artificially patched to produce retransmission and avoid qcs liberation. In order to fix this and avoid future class of similar problem, the best way is to not call qcc_streams_sent_done() to notify MUX for retranmission. To implement this, we test if any of QUIC_FL_CONN_RETRANS_OLD_DATA or the new flag QUIC_FL_CONN_RETRANS_LOST_DATA is set. A new wrapper qc_send_app_retransmit() has been added to set the new flag as a complement to already existing qc_send_app_probing(). This must be backported up to 2.6.	2022-08-17 11:06:24 +02:00
Amaury Denoyelle	cc13047364	MINOR: quic: refactor application send Adjust qc_send_app_pkts function : remove <old_data> arg and provide a new wrapper function qc_send_app_probing() which should be used instead when probing with old data. This simplifies the interface of the default function, most notably for the MUX which does not interfer with retransmission. QUIC_FL_CONN_RETRANS_OLD_DATA flag is set/unset directly in the wrapper qc_send_app_probing(). At the same time, function documentation has been updated to clarified arguments and return values. This commit will be useful for the next patch to differentiate MUX and retransmission send context. As a consequence, the current patch should be backported wherever the next one will be.	2022-08-17 11:05:49 +02:00
Amaury Denoyelle	3baab744e7	MINOR: mux-quic: add missing args on some traces Complete some MUX traces by adding qcc or qcs instance as arguments when this is possible. This will be useful when several connections are interleaved.	2022-08-17 11:05:47 +02:00
Amaury Denoyelle	fd79ddb2d6	MINOR: mux-quic: adjust traces on stream init Adjust traces on qcc_init_stream_remote() : replace "opening" by "initializing" to avoid confusion with traces dealing with OPEN stream state.	2022-08-17 11:05:46 +02:00
Amaury Denoyelle	bf3c208760	BUG/MEDIUM: mux-quic: reject uni stream ID exceeding flow control Emit STREAM_LIMIT_ERROR if a client tries to open an unidirectional stream with an ID greater than the value specified by our flow-control limit. The code is similar to the bidirectional stream opening. MAX_STREAMS_UNI emission is not implement for the moment and is left as a TODO. This should not be too urgent for the moment : in HTTP/3, a client has only a limited use for unidirectional streams (H3 control stream + 2 QPACK streams). This is covered by the value provided by haproxy in transport parameters. This patch has been tagged with BUG as it should have prevented last crash reported on github issue #1808 when opening a new unidirectional streams with an invalid ID. However, it is probably not the main cause of the bug contrary to the patch commit `11a6f4007b` BUG/MINOR: quic: Wrong status returned by qc_pkt_decrypt() This must be backported up to 2.6.	2022-08-17 11:05:19 +02:00
Amaury Denoyelle	26aa399d6b	MINOR: qpack: report error on enc/dec stream close As specified by RFC 9204, encoder and decoder streams must not be closed. If the peer behaves incorrectly and closes one of them, emit a H3_CLOSED_CRITICAL_STREAM connection error. To implement this, QPACK stream decoding API has been slightly adjusted. Firstly, fin parameter is passed to notify about FIN STREAM bit. Secondly, qcs instance is passed via unused void* context. This allows to use qcc_emit_cc_app() function to report a CONNECTION_CLOSE error.	2022-08-17 11:04:53 +02:00
Amaury Denoyelle	6b02c6bb47	MINOR: h3: report error on control stream close As specified by RFC 9114 the control stream must not be closed. If the peer behaves incorrectly and closes it, emit a H3_CLOSED_CRITICAL_STREAM connection error.	2022-08-17 11:04:52 +02:00
Amaury Denoyelle	f372e744de	MINOR: quic: adjust quic_frame flag manipulation Replace a plain '=' operator by '\|=' when setting quic_frame QUIC_FL_TX_FRAME_LOST flag. For the moment, this change has no impact as only two exclusive flags are defined for quic_frame. On the edited code path we are certain that QUIC_FL_TX_FRAME_ACKED is not set due to a previous if statement, so a plain equal or a binary OR is strictly identical. This change will be useful if new flags are defined for quic_frame in the future. These new flags won't be resetted automatically thanks to binary OR without explictly intended, which otherwise could easily lead to new bugs.	2022-08-17 11:04:47 +02:00
Fr�d�ric L�caille	bbeec37b31	MINOR: stick-table: Add table_expire() and table_idle() new converters table_expire() returns the expiration delay for a stick-table entry associated to an input sample. Its counterpart table_idle() returns the time the entry remained idle since the last time it was updated. Both converters may take a default value as second argument which is returned when the entry is not present.	2022-08-17 10:52:15 +02:00
Willy Tarreau	cc1a2a1867	MINOR: chunk: inline alloc_trash_chunk() This function is responsible for all calls to pool_alloc(trash), whose total size can be huge. As such it's quite a pain that it doesn't provide more hints about its users. However, since the function is tiny, it fully makes sense to inline it, the code is less than 0.1% larger with this. This way we can now detect where the callers are via "show profiling", e.g.: 0 1953671 0 32071463136\| 0x59960f main+0x10676f p_free(-16416) [pool=trash] 0 1 0 16416\| 0x59960f main+0x10676f p_free(-16416) [pool=trash] 1953672 0 32071479552 0\| 0x599561 main+0x1066c1 p_alloc(16416) [pool=trash] 0 976835 0 16035723360\| 0x576ca7 http_reply_to_htx+0x447/0x920 p_free(-16416) [pool=trash] 0 1 0 16416\| 0x576ca7 http_reply_to_htx+0x447/0x920 p_free(-16416) [pool=trash] 976835 0 16035723360 0\| 0x576a5d http_reply_to_htx+0x1fd/0x920 p_alloc(16416) [pool=trash] 1 0 16416 0\| 0x576a5d http_reply_to_htx+0x1fd/0x920 p_alloc(16416) [pool=trash]	2022-08-17 10:45:22 +02:00
Willy Tarreau	42b180dcdb	MINOR: pools/memprof: store and report the pool's name in each bin Storing the pointer to the pool along with the stats is quite useful as it allows to report the name. That's what we're doing here. We could store it in place of another field but that's not convenient as it would require to change all functions that manipulate counters. Thus here we store one extra field, as well as some padding because the struct turns 56 bytes long, thus better go to 64 directly. Example of output from "show profiling memory": 2 0 48 0\| 0x4bfb2c ha_quic_set_encryption_secrets+0xcc/0xb5e p_alloc(24) [pool=quic_tls_iv] 0 55252 0 10608384\| 0x4bed32 main+0x2beb2 free(-192) 15 0 2760 0\| 0x4be855 main+0x2b9d5 p_alloc(184) [pool=quic_frame] 1 0 1048 0\| 0x4be266 ha_quic_add_handshake_data+0x2b6/0x66d p_alloc(1048) [pool=quic_crypto] 3 0 552 0\| 0x4be142 ha_quic_add_handshake_data+0x192/0x66d p_alloc(184) [pool=quic_frame] 31276 0 6755616 0\| 0x4bb8f9 quic_sock_fd_iocb+0x689/0x69b p_alloc(216) [pool=quic_dgram] 0 31424 0 6787584\| 0x4bb7f3 quic_sock_fd_iocb+0x583/0x69b p_free(-216) [pool=quic_dgram] 152 0 32832 0\| 0x4bb4d9 quic_sock_fd_iocb+0x269/0x69b p_alloc(216) [pool=quic_dgram]	2022-08-17 10:34:00 +02:00
Willy Tarreau	facfad2b64	MINOR: pool/memprof: report pool alloc/free in memory profiling Pools are being used so well that it becomes difficult to profile their usage via the regular memory profiling. Let's add new entries for pools there, named "p_alloc" and "p_free" that correspond to pool_alloc() and pool_free(). Ideally it would be nice to only report those that fail cache lookups but that's complicated, particularly on the free() path since free lists are released in clusters to the shared pools. It's worth noting that the alloc_tot/free_tot fields can easily be determined by multiplying alloc_calls/free_calls by the pool's size, and could be better used to store a pointer to the pool itself. However it would require significant changes down the code that sorts output. If this were to cause a measurable slowdown, an alternate approach could consist in using a different value of USE_MEMORY_PROFILING to enable pools profiling. Also, this profiler doesn't depend on intercepting regular malloc functions, so we could also imagine enabling it alone or the other one alone or both. Tests show that the CPU overhead on QUIC (which is already an extremely intensive user of pools) jumps from ~7% to ~10%. This is quite acceptable in most deployments.	2022-08-17 09:38:05 +02:00
Willy Tarreau	219afa2ca8	MINOR: memprof: export the minimum definitions for memory profiling Right now it's not possible to feed memory profiling info from outside activity.c, so let's export the function and move the enum and struct to the include file.	2022-08-17 09:03:57 +02:00
Frédéric Lécaille	11a6f4007b	BUG/MINOR: quic: Wrong status returned by qc_pkt_decrypt() This bug came with this big commit: "MEDIUM: quic: xprt traces rework" This is the <ret> variable value which must be returned by most of the xprt functions. This leaded packets which could not be decrypted to be parsed, with weird frames to be parsed as found by Tristan in GH #1808. To be backported where the commit above was backported.	2022-08-16 14:54:32 +02:00
Frédéric Lécaille	ebb1070721	BUG/MINOR: quic: MIssing check when building TX packets When building an ack-eliciting frame only packet, if we did not manage to add at least one such a frame to the packet, we did not notify the caller about the fact the packet is empty. This could lead the caller to believe everything was ok and make it endlessly try to build packet again and again. This issue was amplified by the recent changes where a while(1) loop has been added to qc_send_app_pkt() which calls qc_do_build_pkt() through qc_prep_app_pkts() until we could not prepare packets. Before this recent change, I guess only one empty packet was sent. This patch checks that non empty packets could be built by qc_do_build_pkt() and makes this function return an error if this was the case. Also note that such an issue could happened only when the packet building was limited by the congestion control. Thank you to Tristan for having reported this issue in GH #1808. Must be backported to 2.6.	2022-08-16 12:23:38 +02:00
Amaury Denoyelle	35a66c0a36	BUG/MINOR: mux-quic: fix crash with traces in qc_detach() qc_detach() is used to free a qcs as notified by sedesc. If there is no more stream active and the connection is considered as dead, it will then be freed. This prevent to dereference qcc in TRACE macro. Else this will cause a crash. Use a different code-path on release for qc_detach() to fix this bug. This will fix the last occurence of crash on github issue #1808. This has been introduced by recent QUIC MUX traces rework. Thus, it does not need to be backport.	2022-08-12 16:02:00 +02:00
Willy Tarreau	ded77cc71f	MINOR: ring: archive a previous file-backed ring on startup In order to ensure that an instant restart of the process will not wipe precious debugging information, and to leave time for an admin to archive a copy of a ring, now upon startup, any previously existing file will be renamed with the extra suffix ".bak", and any previously existing file with suffix ".bak" will be removed.	2022-08-12 15:40:19 +02:00
Willy Tarreau	8e87705c21	BUILD: sink: replace S_IRUSR, S_IWUSR with their octal value The build broke on freebsd with S_IRUSR undefined after commit `0b8e9ceb1` ("MINOR: ring: add support for a backing-file"). Maybe another include is needed there, but the point is that we really don't care about these symbolic names, file modes are more readable as 0600 than via these cryptic names anyway, so let's go back to 0600. This will teach me not to try to make things too clean. No backport is needed.	2022-08-12 15:03:12 +02:00
Fr�d�ric L�caille	bfb077acff	BUG/MINOR: quic: memleak on wrong datagram receipt There was a missing pool_free() call for such datagrams. As far as I see there is no leak on valid datagram receipt. Must be backported to 2.6.	2022-08-12 12:19:26 +02:00
Willy Tarreau	0b8e9ceb12	MINOR: ring: add support for a backing-file This mmaps a file which will serve as the backing-store for the ring's contents. The idea is to provide a way to retrieve sensitive information (last logs, debugging traces) even after the process stops and even after a possible crash. Right now this was possible by connecting to the CLI and dumping the contents of the ring live, but this is not handy and consumes quite a bit of resources before it is needed. With a backing file, the ring is effectively RAM-mapped file, so that contents stored there are the same as those found in the file (the OS doesn't guarantee immediate sync but if the process dies it will be OK). Note that doing that on a filesystem backed by a physical device is a bad idea, as it will induce slowdowns at high loads. It's really important that the device is RAM-based. Also, this may have security implications: if the file is corrupted by another process, the storage area could be corrupted, causing haproxy to crash or to overwrite its own memory. As such this should only be used for debugging.	2022-08-12 11:18:46 +02:00
Willy Tarreau	6df10d872b	MINOR: ring: support creating a ring from a linear area Instead of allocating two parts, one for the ring struct itself and one for the storage area, ring_make_from_area() will arrange the two inside the same memory area, with the storage starting immediately after the struct. This will allow to store a complete ring state in shared memory areas for example.	2022-08-12 11:18:46 +02:00
Fr�d�ric L�caille	7629f5d670	BUG/MEDIUM: quic: Wrong use of <token_odcid> in qc_lsntr_pkt_rcv() This commit was not complete: "BUG/MEDIUM: quic: Possible use of uninitialized <odcid> variable in qc_lstnr_params_init()" <token_odcid> should have been directly passed to qc_lstnr_params_init() without dereferencing it to prevent haproxy to have new chances to crash! Must be backported to 2.6.	2022-08-11 19:12:12 +02:00
Willy Tarreau	18d1306abd	BUG/MEDIUM: ring: fix too lax 'size' parser It took me a while to figure why a ring declared with "size 1M" was causing strange effects in a ring, it's just because it's parsed as "1", which is smaller than the default 16384 size and errors are silently ignored. This commit tries to address this the best possible way without breaking existing configs that would work by accident, by warning that the size is ignored if it's smaller than the current one, and by printing the parsed size instead of the input string in warnings and errors. This way if some users have "size 10000" or "size 100k" it will continue to work as 16kB like today but they will now be aware of it. In addition the error messages were a bit poor in context in that they only provided line numbers. The ring name was added to ease locating the problem. As the issue was present since day one and was introduced in 2.2 with commit `99c453df9d` ("MEDIUM: ring: new section ring to declare custom ring buffers."), it could make sense to backport this as far as 2.2, but with 2.2 being quite old now it doesn't seem very reasonable to start emitting new config warnings in config that apparently worked well. Thus it looks more reasonable to backport this as far as 2.4.	2022-08-11 19:05:19 +02:00
Fr�d�ric L�caille	e9325e97c2	BUG/MEDIUM: quic: Possible use of uninitialized <odcid> variable in qc_lstnr_params_init() When receiving a token into a client Initial packet without a cluster secret defined by configuration, the <odcid> variable used to parse the ODCID from the token could be used without having been initialized. Such a packet must be dropped. So the sufficient part of this patch is this check: + } + else if (!global.cluster_secret && token_len) { + /* Impossible case: a token was received without configured + * cluster secret. + */ + TRACE_PROTO("Packet dropped", QUIC_EV_CONN_LPKT, + NULL, NULL, NULL, qv); + goto drop; } Take the opportunity of this patch to rework and make it more readable this part of code where such a packet must be dropped removing the <check_token> variable. When an ODCID is parsed from a token, new <token_odcid> new pointer variable is set to the address of the parsed ODCID. This way, is not set but used it will make crash haproxy. This was not always the case with an uninitialized local variable. Adapt the API to used such a pointer variable: <token> boolean variable is removed from qc_lstnr_params_init() prototype. This must be backported to 2.6.	2022-08-11 18:33:36 +02:00
Amaury Denoyelle	6bdf9367fb	BUG/MEDIUM: mux-quic: fix crash due to invalid trace arg Traces argument were incorrectly used in qcs_free(). A qcs was specified as first arg instead of a connection. This will lead to a crash if developer qmux traces are activated. This is now fixed. This bug has been introduced with QUIC MUX traces rework. No need to backport.	2022-08-11 18:24:53 +02:00
Amaury Denoyelle	4c9a1642c1	MINOR: mux-quic: define new traces Add new traces to help debugging on QUIC MUX. Most notable, the following functions are now traced : * qcc_emit_cc * qcs_free * qcs_consume * qcc_decode_qcs * qcc_emit_cc_app * qcc_install_app_ops * qcc_release_remote_stream * qcc_streams_sent_done * qc_init	2022-08-11 15:20:44 +02:00
Amaury Denoyelle	047d86a34b	CLEANUP: mux-quic: adjust traces level Change default devel level for some traces in QUIC MUX: * proto : used to notify about reception/emission of frames * state : modification of internal state of connection or streams * data : detailled information about transfer and flow-control	2022-08-11 15:20:44 +02:00
Amaury Denoyelle	c7fb0d2b7a	MINOR: mux-quic: define protocol error traces Replace devel traces with error level on all errors situation. Also a new event QMUX_EV_PROTO_ERR is used. This should help to detect invalid situations quickly.	2022-08-11 15:20:44 +02:00
Amaury Denoyelle	f0b67f995c	MINOR: mux-quic: adjust enter/leave traces Improve MUX traces by adding some missing enter/leave trace points. In some places, early function returns have been replaced by a goto statement.	2022-08-11 15:20:43 +02:00
Fr�d�ric L�caille	59507de932	CLEANUP: quic: Remove trailing spaces This spaces have come with this commit: "MEDIUM: quic: xprt traces rework".	2022-08-11 14:33:43 +02:00
Fr�d�ric L�caille	96d08d37d9	BUG/MINOR: quic: Possible infinite loop in quic_build_post_handshake_frames() This loop is due to the fact that we do not select the next node before the conditional "continue" statement. Furthermore the condition and the "continue" statement may be removed after replacing eb64_first() call by eb64_lookup_ge(): we are sure this condition may not be satisfied. Add some comments: this function initializes connection IDs with sequence number 1 upto <max> non included. Take the opportunity of this patch to remove a "return" wich broke this traces rule: for any function, do not call TRACE_ENTER() without TRACE_LEAVE()! Add also TRACE_ERROR() for any encoutered errors. Must be backported to 2.6	2022-08-11 14:33:43 +02:00
Fr�d�ric L�caille	a6920a25d9	MINOR: quic: Remove useless lock for RX packets This lock was there be able to handle the RX packets for a connetion from several threads. This is no more needed since a QUIC connection is always handled by the same thread. May be backported to 2.6	2022-08-11 14:33:43 +02:00
Willy Tarreau	6a378d1677	BUILD: stconn: fix build warning at -O3 about possible null sc gcc-6.x and 7.x emit build warnings about sc possibly being null upon return from sc_detach_endp(). This actually is not the case and the compiler is a little bit overzealous there, but there exists code paths that can make this analysis non-trivial so let's at least add a similar BUG_ON() to let both the compiler and the deverloper know this doesn't happen. This should be backported to 2.6.	2022-08-11 13:59:13 +02:00
Fr�d�ric L�caille	a8b2f843d2	MEDIUM: quic: xprt traces rework Add a least as much as possible TRACE_ENTER() and TRACE_LEAVE() calls to any function. Note that some functions do not have any access to the a quic_conn argument when receiving or parsing datagram at very low level.	2022-08-11 11:11:20 +02:00
Willy Tarreau	e6ca435c04	BUG/MEDIUM: poller: use fd_delete() to release the poller pipes The poller pipes needed to communicate between multiple threads are allocated in init_pollers_per_thread() and released in deinit_pollers_per_thread(). The former adds them via fd_insert() so that they are known, but the former only closes them using a regular close(). This asymmetry represents a problem, because we have in the fdtab[] an entry for something that may disappear when one thread leaves, and since these FD numbers are very low, there is a very high likelihood that they are immediately reassigned to another thread trying to connect() to a server or just sending a health check. In this case, the other thread is going to fd_insert() the fd and the recently added consistency checks will notive that ->owner is not NULL and will crash. We just need to use fd_delete() here to match fd_insert(). Note that this test was added in 2.7-dev2 by commit `36d9097cf` ("MINOR: fd: Add BUG_ON checks on fd_insert()") which was backported to 2.4 as a safety measure (since it allowed to catch particularly serious issues). The patch in itself isn't wrong, it just revealed a long-dormant bug (been there since 1.9-dev1, 4 years ago). As such the current patch needs to be backported wherever the commit above is backported. Many thanks to Christian Ruppert for providing detailed traces in github issue #1807 and Cedric Paillet for bringing his complementary analysis that helped to understand the required conditions for this issue to happen (fast health checks @100ms + randomly long connections ~7s + fast reloads every second + hard-stop-after 5s were necessary on the dev's machine to trigger it from time to time).	2022-08-10 17:25:23 +02:00
Willy Tarreau	54bc78693d	BUG/MEDIUM: quic: always remove the connection from the accept list on close Fred managed to reproduce a crash showing a corrupted accept_list when firing thousands of concurrent picoquicdemo clients to a same instance. It may happen if the connection was placed into the accept_list and immediately closed before being processed (e.g. on error or t/o ?). In any case the quic_conn_release() function should always detach a connection to be deleted from any list, like it does for other lists, so let's add an MT_LIST_DELETE() here. This should be backported to 2.6.	2022-08-10 07:30:22 +02:00
Amaury Denoyelle	f0f92b2db8	BUG/MINOR: quic: fix crash on handshake io-cb for null next enc level When arriving at the handshake completion, next encryption level will be null on quic_conn_io_cb(). Thus this must be check this before dereferencing it via qc_need_sending() to prevent a crash. This was reproduced quickly when browsing over a local nextcloud instance through QUIC with firefox. This has been introduced in the current dev with quic-conn Tx refactoring. No need to backport it.	2022-08-09 18:01:10 +02:00
Amaury Denoyelle	96ca1b7c39	BUG/MINOR: mux-quic: open stream on STOP_SENDING Considered a stream as opened when receiving a STOP_SENDING frame as the first frame on the stream. This patch is tagged as BUG because a BUG_ON may occur if only a STOP_SENDING frame has been received for a frame. This will reset the stream in respect with RFC9000 but internally it is considered invalid transition to reset an idle stream. To fix this, simply use qcs_idle_open() on STOP_SENDING parsing function. This will mark the stream as OPEN before resetting it. This was detected on haproxy.org with the following backtrace : FATAL: bug condition "qcs->st == QC_SS_IDLE" matched at src/mux_quic.c:383 call trace(12): \| 0x490dd3 [b8 01 00 00 00 c6 00 00]: main-0x1d0633 \| 0x4975b8 [48 8b 85 58 ff ff ff 8b]: main-0x1c9e4e \| 0x497df4 [48 8b 45 c8 48 89 c7 e8]: main-0x1c9612 \| 0x49934c [48 8b 45 c8 48 89 c7 e8]: main-0x1c80ba \| 0x6b3475 [48 8b 05 54 1b 3a 00 64]: run_tasks_from_lists+0x45d/0x8b2 \| 0x6b4093 [29 c3 89 d8 89 45 d0 83]: process_runnable_tasks+0x7c9/0x824 \| 0x660bde [8b 05 fc b3 4f 00 83 f8]: run_poll_loop+0x74/0x430 \| 0x6611de [48 8b 05 7b a6 40 00 48]: main-0x228 \| 0x7f66e4fb2ea5 [64 48 89 04 25 30 06 00]: libpthread:+0x7ea5 \| 0x7f66e455ab0d [48 89 c7 e8 5b 72 fc ff]: libc:clone+0x6d/0x86 Stream states have been implemented in the current dev tree. Thus, this patch does not need to be backported.	2022-08-09 17:58:02 +02:00
Amaury Denoyelle	c09ef0c5fc	MINOR: quic: skip sending if no frame to send in io-cb Check on quic_conn_io_cb() if sending is required. This allows to skip over Tx buffer allocation if not needed. To implement this, we check if frame lists on current and next encryption level are empty. We also need to check if there is no need to send ACK, PROBE or CONNECTION_CLOSE. This has been isolated in a new function qc_need_sending() which may be reuse in some other functions in the future.	2022-08-09 16:03:49 +02:00
Amaury Denoyelle	654269c769	MINOR: quic: refactor datagram commit in Tx buffer This is the final patch on quic-conn Tx refactor. Extend the function which is used to write a datagram header to save at the same time written buffer data. This makes sense as the two operations are used at the same occasion when a pre-written datagram is comitted.	2022-08-09 16:00:30 +02:00
Amaury Denoyelle	5b68986d77	MINOR: quic: release Tx buffer on each send Complete refactor of quic-conn Tx buffer. The buffer is now released on every send operation completion. This should help to reduce memory footprint as now Tx buffers are allocated and released on demand. To simplify allocation/free of quic-conn Tx buffer, two static functions are created named qc_txb_alloc() and qc_txb_release().	2022-08-09 16:00:02 +02:00
Amaury Denoyelle	f2476053f9	MINOR: quic: replace custom buf on Tx by default struct buffer On first prototype version of QUIC, emission was multithreaded. To support this, a custom thread-safe ring-buffer has been implemented with qring/cbuf. Now the thread model has been adjusted : a quic-conn is always used on the same thread and emission is not multi-threaded. Thus, qring/cbuf usage can be replace by a standard struct buffer. The code has been simplified even more as for now buffer is always drained after a prepare/send invocation. This is the case since a datagram is always considered as sent even on sendto() error. BUG_ON statements guard are here to ensure that this model is always valid. Thus, code to handle data wrapping and consume too small contiguous space with a 0-length datagram is removed.	2022-08-09 15:45:47 +02:00
Amaury Denoyelle	56c6154dba	CLEANUP: mux-quic: remove loop on sending frames qc_send_app_pkts() has now a while loop implemented which allows to send all possible frames even if the send buffer is full between packet prepare and send. This is present since commit : `dc07751ed7` MINOR: quic: Send packets as much as possible from qc_send_app_pkts() This means we can remove code from the MUX which implement this at the upper layer. This is useful to simplify qc_send_frames() function. As mentionned commit is subject to backport, this commit should be backported as well to 2.6.	2022-08-09 15:41:07 +02:00
Willy Tarreau	4a426e2082	MINOR: debug/memstats: automatically determine first column size The first column's width may vary a lot depending on outputs, and it's annoying to have large empty columns on small names and mangled large columns that are not yet large enough. In order to overcome this, this patch adds a width field to the memstats applet's context, and this width is calculated the first time the function is entered, by estimating the width of all lines that will be dumped. This is simple enough and does the job well. If in the future some filtering criteria are added, it will still be possible to perform a single pass on everything depending on the desired output format.	2022-08-09 08:51:08 +02:00
Willy Tarreau	17200dd1f3	MINOR: debug: also store the function name in struct mem_stats The calling function name is now stored in the structure, and it's reported when the "all" argument is passed. The first column is significantly enlarged because some names are really wide :-(	2022-08-09 08:42:42 +02:00
Willy Tarreau	55c950baa9	MINOR: debug: store and report the pool's name in struct mem_stats Let's add a generic "extra" pointer to the struct mem_stats to store context-specific information. When tracing pool_alloc/pool_free, we can now store a pointer to the pool, which allows to report the pool name on an extra column. This significantly improves tracing capabilities. Example: proxy.c:1598 CALLOC size: 28832 calls: 4 size/call: 7208 dynbuf.c:55 P_FREE size: 32768 calls: 2 size/call: 16384 buffer quic_tls.h:385 P_FREE size: 34008 calls: 1417 size/call: 24 quic_tls_iv quic_tls.h:389 P_FREE size: 34008 calls: 1417 size/call: 24 quic_tls_iv quic_tls.h:554 P_FREE size: 34008 calls: 1417 size/call: 24 quic_tls_iv quic_tls.h:558 P_FREE size: 34008 calls: 1417 size/call: 24 quic_tls_iv quic_tls.h:562 P_FREE size: 34008 calls: 1417 size/call: 24 quic_tls_iv quic_tls.h:401 P_ALLOC size: 34080 calls: 1420 size/call: 24 quic_tls_iv quic_tls.h:403 P_ALLOC size: 34080 calls: 1420 size/call: 24 quic_tls_iv xprt_quic.c:4060 MALLOC size: 45376 calls: 5672 size/call: 8 quic_sock.c:328 P_ALLOC size: 46440 calls: 215 size/call: 216 quic_dgram	2022-08-09 08:26:59 +02:00
Frédéric Lécaille	ba19acd822	MINOR: quic: Replace pool_zalloc() by pool_malloc() for fake datagrams These fake datagrams are only used by the low level I/O handler. They are not provided to the "by connection" datagram handlers. This is why they are not MT_LIST_APPEND()ed to the listner RX buffer list (see &quic_dghdlrs[cid_tid].dgrams in quic_lstnr_dgram_dispatch(). Replace the call to pool_zalloc() to by the lighter call to pool_malloc() and initialize only the ->buf and ->length members. This is safe because only these fields are inspected by the low level I/O handler.	2022-08-08 21:10:58 +02:00
Frédéric Lécaille	ffde3168fc	BUG/MEDIUM: quic: Missing AEAD TAG check after removing header protection After removing the packet header protection, we can check the packet is long enough to contain a 16 bytes length AEAD TAG (at this end of the packet). This test was missing. Must be backported to 2.6.	2022-08-08 18:41:16 +02:00
Frédéric Lécaille	adc7641536	MINOR: quic: Too much useless traces in qc_build_frms() These traces about the available room into the packet currently built and its payload length could be displayed for each STREAM frame, even for those which have no chance to be embedded into a packet leading to very traces to be displayed from a connection with a lot of stream. This was revealed by traces provide by Tristan in GH #1808 May be backported to 2.6.	2022-08-08 16:18:55 +02:00
Frédéric Lécaille	99897d11d9	BUG/MEDIUM: quic: Wrong packet length check in qc_do_rm_hp() When entering this function, we first check the packet length is not too short. But this was done against the datagram lenght in place of the packet length. This could lead to the header protection to be removed using data past the end of the packet (without buffer overflow). Use the packet length in place of the datagram length which is at <end> address passed as parameter to this function. As the packet length is already stored in ->len packet struct member, this <end> parameter is no more useful. Must be backported to 2.6.	2022-08-08 11:02:04 +02:00
Willy Tarreau	2e64472d16	BUILD: cfgparse: always defined _GNU_SOURCE for sched.h and crypt.h _GNU_SOURCE used to be defined only when USE_LIBCRYPT was set. It's also needed for sched_setaffinity() to be exported. As a side effect, when USE_LIBCRYPT is not set, a warning is emitted, as Ilya found and reported in issue #1815. Let's just define _GNU_SOURCE regardless of USE_LIBCRYPT, and also explicitly add sched.h, as right now it appears to be inherited from one of the other includes. This should be backported to 2.4.	2022-08-07 16:55:07 +02:00
Ilya Shipitsin	52f2ff5b93	BUG/MEDIUM: fix DH length when EC key is used dh of length 1024 were chosen for EVP_PKEY_EC key type. let us pick "default_dh_param" instead. issue was found on Ubuntu 22.04 which is shipped with OpenSSL configured with SECLEVEL=2 by default. such SECLEVEL value prohibits DH shorter than 2048: OpenSSL error[0xa00018a] SSL_CTX_set0_tmp_dh_pkey: dh key too small better strategy for chosing DH still may be considered though.	2022-08-06 17:45:40 +02:00
Ilya Shipitsin	3b64a28e15	CLEANUP: assorted typo fixes in the code and comments This is 31st iteration of typo fixes	2022-08-06 17:12:51 +02:00
Willy Tarreau	c80bdb2da6	MINOR: threads: report the number of thread groups in build options haproxy -vv shows the number of threads but didn't report the number of groups, let's add it.	2022-08-06 16:45:26 +02:00
Willy Tarreau	f9d4a7dad3	BUG/MEDIUM: quic: break out of the loop in quic_lstnr_dghdlr The function processes packets sent by other threads in the current thread's queue. But if, for any reason, other threads write faster than the current one processes, this can lead to a situation where the function never returns. It seems that it might be what's happening in issue #1808, though unfortunately, this function is one of the rare without traces. But the amount of calls to functions like qc_lstnr_pkt_rcv() on a single thread seems to indicate this possibility. Thanks to Tristan for his efforts in collecting extremely precious traces! This likely needs to be backported to 2.6.	2022-08-05 16:12:00 +02:00
Amaury Denoyelle	6715cbf97f	BUG/MINOR: quic: adjust errno handling on sendto qc_snd_buf returned a size_t which means that it was never negative despite its documentation. Thus the caller who checked for this was never informed of a sendto error. Clean this by changing the return value of qc_snd_buf() to an integer. A 0 is returned on success. Every other values are considered as an error. This commit should be backported up to 2.6. Note that to not cause malfunctions, it must be backported after the previous patch : `906b058954` MINOR: quic: explicitely ignore sendto error This is to ensure that a sendto error does not cause send to be interrupted which may cause a stalled transfer without a proper retry mechanism. The impact of this bug seems null as caller explicitely ignores sendto error. However this part of code seems to be subject to strange issues and it may fix them in part. It may be of interest for github issue #1808.	2022-08-05 15:53:16 +02:00
Amaury Denoyelle	906b058954	MINOR: quic: explicitely ignore sendto error qc_snd_buf() returns an error if sendto has failed. On standard conditions, we should check for EAGAIN/EWOULDBLOCK errno and if so, register the file-descriptor in the poller to retry the operation later. However, quic_conn uses directly the listener fd which is shared for all QUIC connections of this listener on several threads. Thus, it's complicated to implement fd supversion via the poller : there is no mechanism to easily wakeup quic_conn or MUX after a sendto failure. A quick and simple solution for the moment is to considered a datagram as properly emitted even on sendto error. In the end, this will trigger the quic_conn retransmission timer as data will be considered lost on the network and the send operation will be retried. This solution will be replaced when fd management for quic_conn is reworked. In fact, this quick hack was already in use in the current code, albeit not voluntarily. This is due to a bug caused by an API mismatch on the return type of qc_snd_buf() which never emits a negative error code despite its documentation. Thus, all its invocation were considered as a success. If this bug was fixed, the sending would would have been interrupted by a break which could cause the transfer to freeze. qc_snd_buf() invocation is clean up : the break statement is removed. Send operation is now always explicitely conducted entirely even on error and buffer data is purged. A simple optimization has been added to skip over sendto when looping over several datagrams at the first sendto error. However, to properly function, it requires a fix on the return type of qc_snd_buf() which is provided in another patch. As the behavior before and after this patch seems identical, it is not labelled as a BUG. However, it should be backported for cleaning purpose. It may also have an impact on github issue #1808.	2022-08-05 15:45:25 +02:00
Frédéric Lécaille	e7df68a219	BUG/MINOR: quic: Missing Initial packet dropping case An Initial packet shorter than 1200 bytes must be dropped. The test was there without the "goto drop"! Must be backported to 2.6	2022-08-05 15:27:14 +02:00
Frédéric Lécaille	8ecb7363b5	MINOR: quic: Add two new stats counters for sendto() errors Add "quic_socket_full" new stats counter for sendto() errors with EAGAIN as errno. and "quic_sendto_err" counter for any other error.	2022-08-05 15:27:14 +02:00
Willy Tarreau	af5138fd07	BUG/MINOR: quic: do not reject datagrams matching minimum permitted size The dgram length check in quic_get_dgram_dcid() rejects datagrams matching exactly the minimum allowed length, which doesn't seem correct. I doubt any useful packet would be that small but better fix this to avoid confusing debugging sessions in the future. This might be backported to 2.6.	2022-08-05 10:31:29 +02:00
Willy Tarreau	53bfab080c	BUG/MINOR: sink: fix a race condition between the writer and the reader This is the same issue as just fixed in `b8e0fb97f` ("BUG/MINOR: ring/cli: fix a race condition between the writer and the reader") but this time for sinks. They're also sucking the ring and present the same race at high write loads. This must be backported to 2.2 as well. See comments in the aforementioned commit for backport hints if needed.	2022-08-04 17:21:16 +02:00
Christopher Faulet	96417f392d	BUG/MEDIUM: sink: Set the sink ref for forwarders created during ring parsing A reference to the sink was added in every forwarder by the commit `2ae25ea24` ("MINOR: sink: Add a ref to sink in the sink_forward_target structure"). But this commit is incomplete. It is not performed for the forwarders created during a ring parsing. This patch must be backported to 2.6.	2022-08-04 17:10:28 +02:00
Willy Tarreau	b8e0fb97f3	BUG/MINOR: ring/cli: fix a race condition between the writer and the reader The ring's CLI reader unlocks the read side of a ring and relocks it for writing only if it needs to re-subscribe. But during this time, the writer might have pushed data, see nobody subscribed hence woken nobody, while the reader would have left marking that the applet had no more data. This results in a dump that will not make any forward progress: the ring is clogged by this reader which believes there's no data and the writer will never wake it up. The right approach consists in verifying after re-attaching if the writer had made any progress in between, and to report that another call is needed. Note that a jump back to the beginning would also work but here we provide better fairness between readers this way. This needs to be backported to 2.2. The applet API needed to signal the availability of new data changed a few times since then.	2022-08-04 17:00:21 +02:00
Fr�d�ric L�caille	48bb875908	BUG/MINOR: quic: Avoid sending truncated datagrams There is a remaining loop in this ugly qc_snd_buf() function which could lead haproxy to send truncated UDP datagrams. For now on, we send a complete UDP datagram or nothing! Must be backported to 2.6.	2022-08-03 21:09:04 +02:00
Amaury Denoyelle	30e260e2e6	MEDIUM: mux-quic: implement http-request timeout Implement http-request timeout for QUIC MUX. It is used when the connection is opened and is triggered if no HTTP request is received in time. By HTTP request we mean at least a QUIC stream with a full header section. Then qcs instance is attached to a sedesc and upper layer is then responsible to wait for the rest of the request. This timeout is also used when new QUIC streams are opened during the connection lifetime to wait for full HTTP request on them. As it's possible to demux multiple streams in parallel with QUIC, each waiting stream is registered in a list <opening_list> stored in qcc with <start> as timestamp in qcs for the stream opening. Once a qcs is attached to a sedesc, it is removed from <opening_list>. When refreshing MUX timeout, if <opening_list> is not empty, the first waiting stream is used to set MUX timeout. This is efficient as streams are stored in the list in their creation order so CPU usage is minimal. Also, the size of the list is automatically restricted by flow control limitation so it should not grow too much. Streams are insert in <opening_list> by application protocol layer. This is because only application protocol can differentiate streams for HTTP messaging from internal usage. A function qcs_wait_http_req() has been added to register a request stream by app layer. QUIC MUX can then remove it from the list in qc_attach_sc(). As a side-note, it was necessary to implement attach qcc_app_ops callback on hq-interop module to be able to insert a stream in waiting list. Without this, a BUG_ON statement would be triggered when trying to remove the stream on sedesc attach. This is to ensure that every requests streams are registered for http-request timeout. MUX timeout is explicitely refreshed on MAX_STREAM_DATA and STOP_SENDING frame parsing to schedule http-request timeout if a new stream has been instantiated. It was already done on STREAM parsing due to a previous patch.	2022-08-03 15:04:18 +02:00
Amaury Denoyelle	6ec9837fca	MINOR: mux-quic: refactor refresh timeout function Try to reorganize qcc_refresh_timeout() to improve its readability. The main objective is to reduce the indentation level and if sequences by using goto statement to the end of the function. Also, backend and frontend code path should be more explicit with this new version.	2022-08-03 15:04:18 +02:00
Amaury Denoyelle	418ba21461	MINOR: mux-quic: refresh timeout on frame decoding Refresh the MUX connection timeout in frame parsing functions. This is necessary as these Rx operation are completed directly from the quic-conn layer outside of MUX I/O callback. Thus, the timeout should be refreshed on this occasion. Note that however on STREAM parsing refresh is only conducted when receiving the current consecutive data offset. Timeouts related function have been moved up in the source file to be able to use them in qcc_decode_qcs(). This commit will be useful for http-request timeout. Indeed, a new stream may be opened during qcc_decode_qcs() which should trigger this timeout until a full header section is received and qcs instance is attached to sedesc.	2022-08-03 15:04:18 +02:00
Amaury Denoyelle	8d818c6eab	MINOR: h3: support HTTP request framing state Store the current step of HTTP message in h3s stream. This reports if we are in the parsing of headers, content or trailers section. A new enum h3s_st_req is defined for this. This field is stored in h3s struct but only used for request stream. It is left undefined for other streams (control or QPACK streams). h3_is_frame_valid() has been extended to take into account this state information. A connection error H3_FRAME_UNEXPECTED is reported if an invalid frame according to the current state is received; for example a DATA frame at the beginning of a stream.	2022-08-03 15:04:18 +02:00
Frédéric Lécaille	2c77a5eb8e	BUG/MEDIUM: quic: Floating point exception in cubic_root() It is illegal to call my_flsl() with 0 as parameter value. It is a UB. This leaded cubic_root() to divide values by 0 at this line: x = 2 * x + (uint32_t)(val / ((uint64_t)x * (uint64_t)(x - 1))); Thank you to Tristan971 for having reported this issue in GH #1808 and Willy for having spotted the root cause of this bug. Must follow any cubic for QUIC backport (2.6).	2022-08-03 14:27:20 +02:00
Frédéric Lécaille	8ddde4f05e	BUG/MINOR: quic: Missing in flight ack eliciting packet counter decrement The decrement was missing in quic_pktns_tx_pkts_release() called each time a packet number space is discarded. This is not sure this bug could have an impact during handshakes. This counter is used to cancel the timer used both for packet detection and PTO, setting its value to null. So there could be retransmissions or probing which could be triggered for nothing. Must be backported to 2.6.	2022-08-03 12:59:59 +02:00
Christopher Faulet	6bb86539db	BUG/MEDIUM: proxy: Perform a custom copy for default server settings When a proxy is initialized with the settings of the default proxy, instead of doing a raw copy of the default server settings, a custom copy is now performed by calling srv_settings_copy(). This way, all settings will be really duplicated. Without this deep copy, some pointers are shared between several servers, leading to UAF, double-free or such bugs. This patch relies on following commits: * `b32cb9b51` REORG: server: Export srv_settings_cpy() function * `0b365e3cb` MINOR: server: Constify source server to copy its settings This patch should fix the issue #1804. It must be backported as far as 2.0.	2022-08-03 11:44:34 +02:00
Christopher Faulet	b32cb9b515	REORG: server: Export srv_settings_cpy() function This function will be used to init a proxy with settings of the default proxy. It is mandatory to fix a bug. To do so, it must be exposed.	2022-08-03 11:28:52 +02:00
Christopher Faulet	0b365e3cb5	MINOR: server: Constify source server to copy its settings The source server used to initialize a new server, in srv_settings_cpy() and sub-functions, is now a constant. This patch is mandatory to fix a bug.	2022-08-03 11:28:23 +02:00
Christopher Faulet	bc6b23813f	BUG/MINOR: backend: Don't increment conn_retries counter too early The connection retry counter is incremented too early when a connection fails. In SC_ST_CER state, errors handling must be performed before incrementing the counter. Otherwise, we may consider the max connection attempt is reached while a last one is in fact possible. This patch must be backported to 2.6.	2022-08-03 11:16:35 +02:00
Christopher Faulet	14a60d420a	BUG/MEDIUM: dns: Properly initialize new DNS session When a new DNS session is created, all its fields are not properly initialized. For instance, "tx_msg_offset" can have any value after the allocation. So, to fix the bug, pool_zalloc() is now used to allocate new DNS session. This patch should fix the issue #1781. It must be backported as far as 2.4.	2022-08-03 10:30:07 +02:00
Christopher Faulet	642170a653	BUG/MINOR: peers: Use right channel flag to consider the peer as connected When a peer open a new connection to another peer, it is considered as connected when the hello message is sent. To do so, the peer applet was relying on CF_WRITE_PARTIAL channel flag. However it is not the right flag to use. This one is a transient flag. Depending on the scheduling, this flag may be removed by the stream before the peer has a chance to see it. Instead, CF_WROTE_DATA flag must be checked. This patch is related to the issue #1799. It must be backported as far as 2.0.	2022-08-03 09:56:38 +02:00
Christopher Faulet	160fff665e	BUG/MEDIUM: peers: limit reconnect attempts of the old process on reload When peers are configured and HAProxy is reloaded or restarted, a synchronization is performed between the old process and the new one. To do so, the old process connects on the new one. If the synchronization fails, it retries. However, there is no delay and reconnect attempts are not bounded. Thus, it may loop for a while, consuming all the CPU. Of course, it is unexpected, but it is possible. For instance, if the local peer is misconfigured, an infinite loop can be observed if the connection succeeds but not the synchronization. This prevents the old process to exit, except if "hard-stop-after" option is set. To fix the bug, the reconnect is delayed. The local peer already has a expiration date to delay the reconnects. But it was not used on stopping mode. So we use it not. Thanks to the previous fix, the reconnect timeout is shorter in this case (500ms against 5s on running mode). In addition, we also use the peers resync expiration date to not infinitely retries. It is accurate because the new process, on its side, use this timeout to switch from a local resync to a remote resync. This patch depends on "MINOR: peers: Use a dedicated reconnect timeout when stopping the local peer". It fixes the issue #1799. It should be backported as far as 2.0.	2022-08-03 09:56:38 +02:00
Christopher Faulet	ab4b094055	MINOR: peers: Use a dedicated reconnect timeout when stopping the local peer When a process is stopped or reload, a dedicated reconnect timeout is now used. For now, this timeout is not used because the current code retries immediately to reconnect to perform the local synchronization with the new local peer, if any. This patch is required to fix the issue #1799. It should be backported as far as 2.0 with next fixes.	2022-08-03 09:56:38 +02:00
Christopher Faulet	1b6fa7f5ea	MINOR: peers: Add a warning about incompatible SSL config for the local peer In peers section, it is possible to enable SSL for the local peer. In this case, the bind line and the server line should both be configured. A "default-server" directive may also be used to configure the SSL on the server side. However there is no test to be sure the SSL is enabled on both sides. It is an problem because the local resync performed during a reload will be impossible and it is probably not the expected behavior. So, it is now checked during the configuration validation. A warning message is displayed if the SSL is not properly configured for the local peer. This patch is related to issue #1799. It should probably be backported to 2.6.	2022-08-03 09:56:38 +02:00
Amaury Denoyelle	bd6ec1bf84	MEDIUM: mux-quic: implement http-keep-alive timeout Complete QUIC MUX timeout refresh function by using http-keep-alive timeout. It is used when the connection is idle after having handle at least one request. To implement this a new member <idle_start> has been defined in qcc structure. This is used as timestamp for when the connection became idle and is used as base time for http keep-alive timeout	2022-08-01 15:00:13 +02:00
Amaury Denoyelle	c603de4d84	MINOR: mux-quic: count in-progress requests Add a new qcc member named <nb_hreq>. Its purpose is close to <nb_sc> which represents the number of attached stream connectors. Both are incremented inside qc_attach_sc(). The difference is on the decrement operation. While <nb_cs> is decremented on sedesc detach callback, <nb_hreq> is decremented when the qcs is locally closed. In most cases, <nb_hreq> will be decremented before <nb_cs>. However, it will be the reverse if a stream must be kept alive after detach callback. The main purpose of this field is to implement http-keep-alive timeout. Both <nb_sc> and <nb_hreq> must be null to activate the http-keep-alive timeout.	2022-08-01 14:58:41 +02:00
Amaury Denoyelle	5fc05d17ad	MEDIUM: mux-quic: adjust timeout refresh Implement a new internal function qcc_refresh_timeout(). Its role will be to reset QUIC MUX timeout depending if there is requests in progress or not. qcc_update_timeout() does not set a timeout if there is still attached streams as in this case the upper layer is responsible to manage it. Else it will activate the timeout depending on the connection current status. Timeout is refreshed on several locations : on stream detach and in I/O handler and wake callback. For the moment, only the default timeout is used (client or server). The function may be expanded in the future to support more specific ones : * http-keep-alive if connection is idle * http-request when waiting for incomplete HTTP requests * client/server-fin for graceful shutdown	2022-08-01 14:58:36 +02:00
Amaury Denoyelle	b6309456d0	MINOR: mux-quic: use timeout server for backend conns Use timeout server in qcc_init() as default timeout for backend connections. No impact for the moment as QUIC backend support is not implemented.	2022-08-01 14:23:21 +02:00
Amaury Denoyelle	07bf8f4d86	MINOR: mux-quic: save proxy instance into qcc Store a reference to proxy in the qcc structure. This will be useful to access to proxy members outside of qcc_init(). Most notably, this change is required to implement timeout refreshing by using the various timeouts configured at the proxy level.	2022-08-01 14:23:21 +02:00
Amaury Denoyelle	09ec3e09bd	BUG/MINOR: mux-quic: do not free conn if attached streams Ensure via qcc_is_dead() that a connection is not released instance until all of qcs streams are detached by the upper layer, even if an error has been reported or the timeout has fired. On the other side, as qc_detach() always check the connection status, this should ensure that we do not keep a connection if not necessary. Without this patch, a qcc instance may be freed with some of its qcs streams not detached. This is an incorrect behavior and will lead to a BUG_ON fault. Note however that no occurence of this bug has been produced currently. This patch is mainly a safety against future occurences. This should be backported up to 2.6.	2022-08-01 14:23:19 +02:00
Amaury Denoyelle	4ea5090f55	CLEANUP: mux-quic: remove useless app_ops is_active callback Timeout in QUIC MUX has evolved from the simple first implementation. At the beginning, a connection was considered dead unless bidirectional streams were opened. This was abstracted through an app callback is_active(). Now this paradigm has been reversed and a connection is considered alive by default, unless an error has been reported or a timeout has already been fired. The callback is_active() is thus not used anymore and can be safely removed to simplify qcc_is_dead(). This commit should be backported to 2.6.	2022-08-01 14:13:51 +02:00
Amaury Denoyelle	d3973853c2	BUG/MINOR: mux-quic: prevent crash if conn released during IO callback A qcc instance may be freed in the middle of qc_io_cb() if all streams were purged. This will lead to a crash as qcc instance is reused after this step. Jump directly to the end of the function to avoid this. Note that this bug has not been triggered for the moment. This is a safety fix to prevent it. This must be backported up to 2.6.	2022-08-01 14:13:51 +02:00
Willy Tarreau	51d38a26fe	BUG/MEDIUM: pattern: only visit equivalent nodes when skipping versions Miroslav reported in issue #1802 a problem that affects atomic map/acl updates. During an update, incorrect versions are properly skipped, but in order to do so, we rely on ebmb_next() instead of ebmb_next_dup(). This means that if a new matching entry is in the process of being added and is the first one to succeed in the lookup, we'll skip it due to its version and use the next entry regardless of its value provided that it has the correct version. For IP addresses and string prefixes it's particularly visible because a lookup may match a new longer prefix that's not yet committed (e.g. 11.0.0.1 would match 11/8 when 10/7 was the only committed one), and skipping it could end up on 12/8 for example. As soon as a commit for the last version happens, the issue disappears. This problem only affects tree-based matches: the "str", "ip", and "beg" matches. Here we replace the ebmb_next() values with ebmb_next_dup() for exact string matches, and with ebmb_lookup_shorter() for longest matches, which will first visit duplicates, then look for shorter prefixes. This relies on previous commit: MINOR: ebtree: add ebmb_lookup_shorter() to pursue lookups Both need to be backported to 2.4, where the generation ID was added. Note that nowadays a simpler and more efficient approach might be employed, by having a single version in the current tree, and a list of trees per version. Manipulations would look up the tree version and work (and lock) only in the relevant trees, while normal operations would be performed on the current tree only. Committing would just be a matter of swapping tree roots and deleting old trees contents.	2022-08-01 11:59:46 +02:00
Willy Tarreau	0dc9e6dca2	DEBUG: tools: provide a tree dump function for ebmbtrees as well It's convenient for debugging IP trees. However we're not dumping the full keys, for the sake of simplicity, only the 4 first bytes are dumped as a u32 hex value. In practice this is sufficient for debugging. As a reminder since it seems difficult to recover the command each time it's needed, the output is converted to an image using dot from Graphviz: dot -o a.png -Tpng dump.txt	2022-08-01 11:59:15 +02:00
Willy Tarreau	87aff021db	MINOR: thread: provide an alternative to pthread's rwlock Since version 1.1.0, OpenSSL's libcrypto ignores the provided locking mechanism and uses pthread's rwlocks instead. The problem is that for some code paths (e.g. async engines) this results in a huge amount of syscalls on systems facing a bit of contention, to the point where more than 80% of the CPU can be spent in the system dealing with spinlocks just for futex_wake(). This patch provides an alternative by redefining the relevant pthread rwlocks from the low-overhead version of the progressive rw locks. This way there will be no more syscalls in case of contention, and CPU will be burnt in userland. Doing this saves massive amounts of CPU, where the locks only take 12-15% vs 80% before, which allows SSL to work much faster on large thread counts (e.g. 24 or more). The tryrdlock and trywrlock variants have been implemented using a CAS since their goal is only to succeed on no contention and never to wait. The pthread_rwlock API is complete except that the timed versions of the rdlock and wrlock do not wait and simply fall back to trylock versions. Since the gains have only been observed with async engines for now, this option remains disabled by default. It can be enabled at build time using USE_PTHREAD_EMULATION=1.	2022-07-30 10:17:22 +02:00
Willy Tarreau	ddab05b98a	BUG/MEDIUM: queue/threads: limit the number of entries dequeued at once When testing strong queue contention on a 48-thread machine, some crashes would frequently happen due to process_srv_queue() never leaving and processing pending requests forever. A dump showed more than 500000 loops at once. The problem is that other threads find it working so they don't do anything and are free to process their pending requests. Because of this, the dequeuing thread can be kept busy forever and does not process its own requests anymore (fortunately the watchdog stops it). This patch adds a limit to the number of rounds, it limits it to maxpollevents, which is reasonable because it's also an indicator of latency and batches size. However there's a catch. If all requests are finished when the thread ends the loop, there might not be enough anymore to restart processing the queue. Thus we tolerate to re-enter the loop to process one request at a time when it doesn't have any anymore. This way we're leaving more room for another thread to take on this task, and we're sure to eventually end this loop. Doing this has also improved the overall dequeuing performance by ~20% in highly contended situations with 48 threads. It should be backported at least to 2.4, maybe even 2.2 since issues were faced in the past on machines having many cores.	2022-07-30 10:00:59 +02:00
Frédéric Lécaille	dc07751ed7	MINOR: quic: Send packets as much as possible from qc_send_app_pkts() Add a loop into this function to send more packets from this function which is called by the mux. It is broken when we could not prepare packet with qc_prep_app_pkts() due to missing available room in the buffer used to send packets. This improves the throughput. Must be backported to 2.6.	2022-07-29 17:32:05 +02:00
Frédéric Lécaille	843399fd45	BUG/MAJOR: quic: Useless resource intensive loop qc_ackrng_pkts() This usless loop should have been removed a long time ago. As it is CPU resource intensive, it could trigger the watchdog. Must be backported to 2.6.	2022-07-29 17:32:05 +02:00
Frédéric Lécaille	dc591cd6cb	MINOR: quic: Stop looking for packet loss asap As the TX packets are ordered by their packet number and always sent in the same order. their TX timestamps are inspected from the older to the newer values when we look for the packet loss. So we can stop this search as soon as we found the first packet which has not been lost. Must be backported to 2.6	2022-07-29 17:32:05 +02:00
Frédéric Lécaille	d2e104ff78	BUG/MINOR: quic: loss time limit variable computed but not used <loss_time_limit> is the loss time limit computed from <time_sent> packet transmission timestamps in qc_packet_loss_lookup() to identify the packets which have been lost. This latter timestamp variable was used in place of <loss_time_limit> to distinguish such packets from others (still in fly packets). Must be backported to 2.6	2022-07-29 17:32:05 +02:00
Frédéric Lécaille	43910a9450	MINOR: quic: New "quic-cc-algo" bind keyword As it could be interesting to be able to choose the QUIC control congestion algorithm to be used by listener, add "quic-cc-algo" new keyword to do so. Update the documentation consequently. Must be backported to 2.6.	2022-07-29 17:32:05 +02:00
Frédéric Lécaille	1c9c2f6c02	MEDIUM: quic: Cubic congestion control algorithm implementation Cubic is the congestion control algorithm used by default by the Linux kernel since 2.6.15 version. This algorithm is supposed to achieve good scalability and fairness between flows using the same network path, it should also be used by QUIC by default. This patch implements this algorithm and select it as default algorithm for the congestion control. Must be backported to 2.6.	2022-07-29 17:32:05 +02:00
Frédéric Lécaille	c591459d11	MINOR: quic: Congestion control architecture refactoring Ease the integration of new congestion control algorithm to come. Move the congestion controller state to a private array of uint32_t to stop using a union. We do not want to continue using such long paths cc->algo_state.<algo>.<var> to modify the internal state variable for each algorithm. Must be backported to 2.6	2022-07-29 17:32:05 +02:00
Amaury Denoyelle	72a78e8290	BUG/MEDIUM: mux-quic: fix missing EOI flag to prevent streams leaks On H3 DATA frame transfer from the client, some streams are not properly closed by the upper layer, despite all transfer operation completed. Data integrity is not impacted but this will prevent the stream timeout to fire and thus keep the owner session opened. In most cases, sessions are closed on QUIC idle timeout, but it may stay forever if a client emits PING frames at a regular interval to maintain it. This bug is caused by a missing EOI stream desc flag on certain condition in qc_rcv_buf(). To be triggered, we have to use the optimization when conn-stream buffer is empty and can be swapped with qcs buffer. The problem is that it will skip the function body for default copy but also a condition to check if EOI must be set. Thus this bug does not happens for every H3 post requets : it requires that conn-stream buffer is empty on last qc_rcv_buf() invocation. This was reproduced more frequently when using ngtcp2 client with one or multiple streams : $ ngtcp2-client -m POST -d ~/infra/html/10K 127.0.0.1 20443 \ http://127.0.0.1:20443/post This may fix at least partially github issue #1801. This must be backported up to 2.6.	2022-07-29 16:01:21 +02:00
William Lallemand	b5d062dff1	MINOR: cli: warning on _getsocks when socket were closed The previous attempt was reverted because it would emit a warning when the sockets are still in the process when a reload failed, so this was an expected 2nd try. This warning however, will be displayed if a new process successfully get the previous sockets AND the sendable number of sockets is 0. This way the user will be warned if he tried to get the sockets fromt the wrong process.	2022-07-28 15:49:43 +02:00
William Lallemand	9c821e615e	Revert "MINOR: cli: emit a warning when _getsocks was used more than once" This reverts commit `519cd2021b`. This was reverted because it's still useful to have access to _getsosks when the previous reload failed.	2022-07-27 13:55:54 +02:00
William Lallemand	14b98ef1bd	BUG/MINOR: mworker: PROC_O_LEAVING used but not updated Since commit `2be557f` ("MEDIUM: mworker: seamless reload use the internal sockpair"), we are using the PROC_O_LEAVING flag to determine which sockpair worker will be used with -x during the next reload. However in mworker_reexec(), the PROC_O_LEAVING flag is not updated, it is only updated at startup in mworker_env_to_proc_list(). This could be a problem when a remaining process is still in the list, it could be selected as the current worker, and its socket will be used even if _getsocks doesn't work anymore on it. (bug #1803) This patch fixes the issue by updating the PROC_O_LEAVING flag in mworker_proc_list_to_env() just before using it in mworker_reexec() Must be backported to 2.6.	2022-07-27 12:13:56 +02:00
William Lallemand	519cd2021b	MINOR: cli: emit a warning when _getsocks was used more than once The _getsocks CLI command can be used only once, after that the sockets are not available anymore. Emit a warning when the command was already used once.	2022-07-27 11:48:54 +02:00
Willy Tarreau	b983145837	BUG/MINOR: fd: always remove late updates when freeing fd_updt[] Christopher found that since commit `8e2c0fa8e` ("MINOR: fd: delete unused updates on close()") we may crash in a late stop due to an fd_delete() in the main thread performed after all threads have deleted the fd_updt[] array. Prior to that commit that didn't happen because we didn't touch the updates on this path, but now it may happen. We don't care about these ones anyway since the poller is stopped, so let's just wipe them by resetting their counter before freeing the array. No backport is needed as this is only 2.7.	2022-07-26 19:06:17 +02:00
William Lallemand	c31577f32e	MEDIUM: resolvers: continue startup if network is unavailable When haproxy starts with a resolver section, and there is a default one since 2.6 which use /etc/resolv.conf, it tries to do a connect() with the UDP socket in order to check if the routes of the system allows to reach the server. This check is too much restrictive as it won't prevent any runtime failure. Relax the check by making it a warning instead of a fatal alert. This must be backported in 2.6.	2022-07-26 10:59:14 +02:00
Christopher Faulet	244331f6e7	Revert "BUG/MINOR: peers: set the proxy's name to the peers section name" This reverts commit `356866acce`. It seems that an undocumented expectation of peers is based on the peers proxy name to determine if the local peer is fully configured or not. Thus because of the commit above, we are no longer able to detect incomplete peers sections. On side effect of this bug is a segfault when HAProxy is stopped/reloaded if we try to perform a local resync on a mis-configured local peer. So waiting for a better solution, the patch is reverted. This patch must be backported as far as 2.5.	2022-07-25 16:17:04 +02:00
William Lallemand	708949da49	MINOR: sockpair: move send_fd_uxst() error message in caller Move the ha_alert() in send_fd_uxst() in the callers and add the FD numbers in the message.	2022-07-25 16:11:11 +02:00
William Lallemand	f67e8fb92c	BUG/MINOR: sockpair: wrong return value for fd_send_uxst() The fd_send_uxst() function which is used to send a socket over the socketpair returns 1 upon error instead of -1, which means the error case of the sendmsg() is never catched correctly. Must be backported as far as 1.9.	2022-07-25 16:10:58 +02:00
Willy Tarreau	6983426354	BUG/MAJOR: poller: drop FD's tgid when masks don't match A bug was introduced in 2.7-dev2 by commit `1f947cb39` ("MAJOR: poller: only touch/inspect the update_mask under tgid protection"): once the FD's tgid is held, we would forget to drop it in case the update mask doesn't match, resulting in random watchdog panics of older processes on successive reloads. This should fix issue #1798. Thanks to Christian for the report and to Christopher for the reproducer. No backport is needed.	2022-07-25 15:47:15 +02:00
Willy Tarreau	53bfac8c63	BUG/MEDIUM: master: force the thread count earlier Christopher bisected that recent commit `d0b73bca71` ("MEDIUM: listener: switch bind_thread from global to group-local") broke the master socket in that only the first out of the Nth initial connections would work, where N is the number of threads, after which they all work. The cause is that the master socket was bound to multiple threads, despite global.nbthread being 1 there, so the incoming connection load balancing would try to send incoming connections to non-existing threads, however the bind_thread mask would nonetheless include multiple threads. What happened is that in 1.9 we forced "nbthread" to 1 in the master's poll loop with commit `b3f2be338b` ("MEDIUM: mworker: use the haproxy poll loop"). In 2.0, nbthread detection was enabled by default in commit `149ab779cc` ("MAJOR: threads: enable one thread per CPU by default"). From this point on, the operation above is unsafe because everything during startup is performed with nbthread corresponding to the default value, then it changes to one when starting the polling loop. But by then we weren't using the wait mode except for reload errors, so even if it would have happened nobody would have noticed. In 2.5 with commit `fab0fdce9` ("MEDIUM: mworker: reexec in waitpid mode after successful loading") we started to rexecute all the time, not just for errors, so as to release precious resources and to possibly spot bugs that were rarely exposed in this mode. By then the incoming connection LB was enforcing all_threads_mask on the listener's thread mask so that the incorrect value was being corrected while using it. Finally in 2.7 commit `d0b73bca71` ("MEDIUM: listener: switch bind_thread from global to group-local") replaces the all_threads_mask there with the listener's bind_thread, but that one was never adjusted by the starting master, whose thread group was filled to N threads by the automatic detection during early setup. The best approach here is to set nbthread to 1 very early in init() when we're in the master in wait mode, so that we don't try to guess the best value and don't end up with incorrect bindings anymore. This patch does this and also sets nbtgroups to 1 in preparation for a possible future where this will also be automatically calculated. There is no need to backport this patch since no other versions were affected, but if it were to be discovered that the incorrect bind mask on some of the master's FDs could be responsible for any trouble in older versions, then the backport should be safe (provided that nbtgroups is dropped of course).	2022-07-22 17:51:53 +02:00
Christopher Faulet	38c53944cb	BUG/MINOR: backend: Fallback on RR algo if balance on source is impossible If the loadbalancing is performed on the source IP address, an internal error was returned on error. So for an applet on the client side (for instance an SPOE applet) or for a client connected to a unix socket, an internal error is returned. However, when other LB algos fail, a fallback on round-robin is performed. There is no reson to not do the same here. This patch should fix the issue #1797. It must be backported to all supported versions.	2022-07-22 17:07:34 +02:00
Christopher Faulet	ca67992979	BUG/MEDIUM: stconn: Only reset connect expiration when processing backend side Since commit `ae024ced0` ("MEDIUM: stream-int/stream: Use connect expiration instead of SI expiration"), the connect expiration date is per-stream. So there is only one expiration date instead of one per side, front and back. So when a stream-connector is processed, we must test if it is a frontend or a backend stconn before updating the connect expiration date. Indeed, the frontend stconn must not reset the connect expiration date. This bug may have several side effect. One known bug is about peer sessions blocked because the frontend peer applet is in ST_CLO state and its backend connection is in ST_TAR state but without connect expiration date. This patch should fix the issue #1791 and #1792. It must be backported to 2.6.	2022-07-21 14:50:14 +02:00
Willy Tarreau	41afd9084e	BUILD: add detection for unsupported compiler models As reported in github issue #1765, some people get trapped into building haproxy and companion libraries on Windows using a compiler following the LLP64 model. This has no chance to work, and definitely causes nasty bugs everywhere when pointers are passed as longs. Let's save them time and detect this at boot time. The message and detection was factored with the existing one for -fwrapv since we need the same info and actions. This should be backported to all recent supported versions (the ones that are likely to be tried on such platforms when people don't know).	2022-07-21 09:58:20 +02:00
William Lallemand	d4835a9680	BUG/MEDIUM: mworker: proc_self incorrectly set crashes upon reload When updating from 2.4 to 2.6, the child->reloads++ instruction changed place, resulting in a former worker from the 2.4 process, still identified as a current worker once in 2.6, because its reload counter is still 0. Unfortunately this counter is used to chose the mworker_proc structure that will be used for the new worker. What happens next, is that the mworker_proc structure of the previous process is selected, and this one has ipc_fd[1] set to -1, because this structure was supposed to be in the master. The process then forks, and mworker_sockpair_register_per_thread() tries to register ipc_fd[1] which is set to -1, instead of the fd of the new socketpair. This patch fixes the issue by checking if child->pid is equal to -1 when selecting proc_self. This way we could be sure it wasn't a previous process. Should fix issue #1785. This must be backported as far as 2.4 to fix the issue related to the reload computation difference. However backporting it in every stable branch will enforce the reload process.	2022-07-21 00:52:43 +02:00
Frédéric Lécaille	a18c3339c8	BUG/MAJOR: mux_quic: fix invalid PROTOCOL_VIOLATION on POST data overlap Stream data reception is incorrect when dealing with a partially new offset with some data already consumed out of the RX buffer. In this case, data length is adjusted but not the data buffer. In most cases, ncb_add() operation will be rejected as already stored data does not correspond with the new inserted offset. This will result in an invalid CONNECTION_CLOSE with PROTOCOL_VIOLATION. To fix this, buffer pointer is advanced while the length is reduced. This can be reproduced with a POST request and patching haproxy to call qcc_recv() multiple times by copying a quic_stream frame with different offsets. Must be backported to 2.6.	2022-07-20 15:34:58 +02:00
William Lallemand	bac3a82a50	BUG/MINOR: mworker/cli: relative pid prefix not validated anymore Since `e8422bf` ("MEDIUM: global: remove the relative_pid from global and mworker"), the relative pid prefix is not tested anymore on the master CLI. Which means any value will fall into the "1" process. Since we removed the nbproc, only the "1" and the "0" (master) value are correct, any other value should return an error. Fix issue #1793. This must be backported as far as 2.5.	2022-07-20 14:43:47 +02:00
William Lallemand	0f17ab2fdd	MINOR: ssl: enhance ca-file error emitting Enhance the errors and warnings when trying to load a ca-file with ssl_store_load_locations_file(). Add errors from ERR_get_error() and strerror to give more information to the user.	2022-07-19 19:13:08 +02:00
William Lallemand	3b8bafd4a7	MINOR: init: load OpenSSL error strings Load OpenSSL Error strings in order to be able to output reason strings. This is mandatory to be able to use ERR_reason_error_string().	2022-07-19 19:13:08 +02:00
Willy Tarreau	c1640f79fe	BUG/MEDIUM: fd/threads: fix incorrect thread selection in wakeup broadcast In commit `cfdd20a0b` ("MEDIUM: fd: support broadcasting updates for foreign groups in updt_fd_polling") we decided to pick a random thread number among a set of candidates for a wakeup in case we need an instant change. But the thread count range was wrong (MAX_THREADS) instead of tg->count, resulting in random crashes when thread groups are > 1 and MAX_THREADS > 64. No backport is needed, this was introduced in 2.7-dev2.	2022-07-19 16:01:04 +02:00
Christopher Faulet	f7ebe584d7	BUILD: debug: Add braces to if statement calling only CHECK_IF() In src/ev_epoll.c, a CHECK_IF() is guarded by an if statement. So, when the macro is empty, GCC (at least 11.3.1) is not happy because there is an if statement with an empty body without braces... It is handled by "-Wempty-body" option. So, braces are added and GCC is now happy. No backport needed.	2022-07-19 12:11:04 +02:00
Amaury Denoyelle	0933c7b3c8	BUG/MINOR: quic: do not send CONNECTION_CLOSE_APP in initial/handshake As specified by RFC 9000, it is forbidden to send a CONNECTION_CLOSE of type 0x1d (CONNECTION_CLOSE_APP) in an Initial or Handshake packet. It must be converted to type 0x1c (CONNECTION_CLOSE) with APPLICATION_ERROR code. CONNECTION_CLOSE_APP are generated by QUIC MUX interaction. Thus, special care must be taken when dealing with a 0-RTT packet, as this is the only case where the MUX can be instantiated and quic-conn still on the Initial or Handshake encryption level. To enforce RFC 9000, xprt build packet function is now responsible to translate a CONNECTION_CLOSE_APP if still on Initial/Handshake encryption. This process is done in a dedicated function named qc_build_cc_frm(). Without this patch, BUG_ON() statement in qc_build_frm() will be triggered when building a CONNECTION_CLOSE_APP frame on Initial or Handshake level. This is because QUIC_FT_CONNECTION_CLOSE_APP frame builder mask does not allow these encryption levels, as opposed to QUIC_FT_CONNECTION_CLOSE builder. This crash was reproduced by modifying the H3 layer to force emission of a CONNECTION_CLOSE_APP on first frame of a 0-RTT session. Note however that CONNECTION_CLOSE emission during Handshake is a complicated process for the server. For the moment, this is still incomplete on haproxy side. RFC 9000 requires to emit it multiple times in several packets under different encryption levels, depending on what we know about the client encryption context. This patch should be backported up to 2.6.	2022-07-19 11:19:50 +02:00
William Lallemand	4348232231	BUG/MINOR: ssl: allow duplicate certificates in ca-file directories It looks like OpenSSL 1.0.2 returns an error when trying to insert a certificate whis is already present in a X509_STORE. This patch simply ignores the X509_R_CERT_ALREADY_IN_HASH_TABLE error if emitted. Should fix part of issue #1780. Must be backported in 2.6.	2022-07-18 18:49:27 +02:00
William Lallemand	3bda80789c	BUG/MINOR: resolvers: shut off the warning for the default resolvers When the resolv.conf file is empty or there is no resolv.conf file, an empty resolvers will be created, which emits a warning during the postparsing step. This patch fixes the problem by freeing the resolvers section if the parsing failed or if the nameserver list is empty. Must be backported in 2.6, the previous patch which introduces resolvers_destroy() is also required.	2022-07-18 14:39:36 +02:00
William Lallemand	e606c84fee	MINOR: resolvers: resolvers_destroy() deinit and free a resolver Split the resolvers_deinit() function into resolvers_destroy() and resolvers_deinit() in order to be able to free a unique resolvers section.	2022-07-18 14:39:36 +02:00
Willy Tarreau	5b3cd9561b	BUG/MEDIUM: tools: avoid calling dlsym() in static builds (try 2) The first approach in commit `288dc1d8e` ("BUG/MEDIUM: tools: avoid calling dlsym() in static builds") relied on dlopen() but on certain configs (at least gcc-4.8+ld-2.27+glibc-2.17) it used to catch situations where it ought not fail. Let's have a second try on this using dladdr() instead. The variable was renamed "build_is_static" as it's exactly what's being detected there. We could even take it for reporting in -vv though that doesn't seem very useful. At least the variable was made global to ease inspection via the debugger, or in case it's useful later. Now it properly detects a static build even with gcc-4.4+glibc-2.11.1 and doesn't crash anymore.	2022-07-18 14:03:54 +02:00
Willy Tarreau	288dc1d8ee	BUG/MEDIUM: tools: avoid calling dlsym() in static builds Since 2.4 with commit `64192392c` ("MINOR: tools: add functions to retrieve the address of a symbol"), we can resolve symbols. However some old glibc crash in dlsym() when the program is statically built. Fortunately even on these old libs we can detect lack of support by calling dlopen(NULL). Normally it returns a handle to the current program, but on a static build it returns NULL. This is sufficient to refrain from calling dlsym() (which will be of very limited use anyway), so we check this once at boot and use the result when needed. This may be backported to 2.4. On stable versions, be careful to place the init code inside an if/endif guard that checks for DL support.	2022-07-16 13:49:34 +02:00
Willy Tarreau	c6b596dcce	CLEANUP: threads: remove the now unused all_threads_mask and tid_bit Since these are not used anymore, let's now remove them. Given the number of places where we're using ti->ldit_bit, maybe an equivalent might be useful though.	2022-07-15 20:25:41 +02:00
Willy Tarreau	cfdd20a0b2	MEDIUM: fd: support broadcasting updates for foreign groups in updt_fd_polling We're still facing the situation where it's impossible to update an FD for a foreign group. That's of particular concern when disabling/enabling listeners (e.g. pause/resume on signals) since we don't decide which thread gets the signal and it needs to process all listeners at once. Fortunately, not that much is unprotected in FDs. This patch adds a test for tgid's equality in updt_fd_polling() so that if a change is applied for a foreing group, then it's detected and taken care of separately. The method consists in forcing the update on all bound threads in this group, adding it to the group's update_list, and sending a wake-up as would be done for a remote thread in the local group, except that this is done by grabbing a reference to the FD's tgid. Thanks to this, SIGTTOU/SIGTTIN now work for nbtgroups > 1 (after that was temporarily broken by "MEDIUM: fd/poller: make the update-list per-group").	2022-07-15 20:25:41 +02:00
Willy Tarreau	1f947cb39e	MAJOR: poller: only touch/inspect the update_mask under tgid protection With thread groups and group-local masks, the update_mask cannot be touched nor even checked if it may change below us. In order to avoid this, we have to grab a reference to the FD's tgid before checking the update mask. The operations are cheap enough so that we don't notice it in performance tests. This is expected because the risk of meeting a reassigned FD during an update remains very low. It's worth noting that the tgid cannot be trusted during startup nor during soft-stop since that may come from anywhere at the moment. Since soft-stop runs under thread isolation we use that hint to decide whether or not to check that the FD's tgid matches the current one. The modification is applied to the 3 thread-aware pollers, i.e. epoll, kqueue, and evports. Also one poll_drop counter was missing for shared updates, though it might be hard to trigger it. With this change applied, thread groups are usable in benchmarks.	2022-07-15 20:16:30 +02:00
Willy Tarreau	d95f18fa39	MAJOR: pollers: rely on fd_reregister_all() at boot time The poller-specific thread init code now uses that new function to safely register boot events. This ensures that we don't register an event for another group and that we properly deal with parallel thread startup. It's only done for thread-aware pollers, there's no point in using that in poll/select though that should work as well.	2022-07-15 20:16:30 +02:00
Willy Tarreau	9baff4ffd9	MEDIUM: fd: support stopping FDs during starting There's a nasty case during boot, which is the master process. It stops all listeners from the main thread, and as such we're seeing calls to fd_delete() from a thread that doesn't match the FD's mask, but more importantly from a group that doesn't match either. Fortunately this happens in a process that doesn't see the threads creation, so the FDs are left intact in the table and we can overwrite the tgid there. The approach is ugly, it probably shows that we should use a dummy value for the tgid during boot, that would be replaced once the FDs migrate to their target, but we also need a way to make sure not to miss them. Also that doesn't solve the possibility of closing a listener at run time from the wrong thread group.	2022-07-15 20:16:30 +02:00
Willy Tarreau	88c4c14050	MINOR: fd: add fd_reregister_all() to deal with boot-time FDs At boot the pollers are allocated for each thread and they need to reprogram updates for all FDs they will manage. This code is not trivial, especially when trying to respect thread groups, so we'd rather avoid duplicating it. Let's centralize this into fd.c with this function. It avoids closed FDs, those whose thread mask doesn't match the requested one or whose thread group doesn't match the requested one, and performs the update if required under thread-group protection.	2022-07-15 20:16:30 +02:00
Willy Tarreau	d0b73bca71	MEDIUM: listener: switch bind_thread from global to group-local It requires to both adapt the parser and change the algorithm to redispatch incoming traffic so that local threads IDs may always be used. The internal structures now only reference thread group IDs and group-local masks which are compatible with those now used by the FD layer and the rest of the code.	2022-07-15 20:16:30 +02:00
Willy Tarreau	6018c02c36	MEDIUM: thread: change thread_resolve_group_mask() to return group-local values It used to turn group+local to global but now we're doing the exact opposite as we want to stick to group-local masks. This means that "thread 3-4" might very well emit what "thread 2/1-2" used to emit till now for 2 groups and 4 threads. This is needed because we'll have to support group-local thread masks in receivers. However the rest of the code (receivers) is not ready yet for this, so using this code with more than one thread group will definitely break some bindings.	2022-07-15 20:16:30 +02:00
Willy Tarreau	0b51eab764	MEDIUM: fd: quit fd_update_events() when FD is closed The IOCB might have closed the FD itself, so it's not an error to have fd.tgid==0 or anything else, nor to have a null running_mask. In fact there are different conditions under which we can leave the IOCB, all of them have been enumerated in the code's comments (namely FD still valid and used, hence has running bit, FD closed but not yet reassigned thus running==0, FD closed and reassigned, hence different tgid and running becomes irrelevant, just like all other masks). For this reason we have no other solution but to try to grab the tgid on return before checking the other bits. In practice it doesn't represent a big cost, because if the FD was closed and reassigned, it's instantly detected and the bit is immediately released without blocking other threads, and if the FD wasn't closed this doesn't prevent it from being migrated to another thread. In the worst case a close by another thread after a migration will be postponed till the moment the running bit is cleared, which is the same as before.	2022-07-15 20:16:30 +02:00
Willy Tarreau	ddedc16624	MEDIUM: fd: make fd_insert/fd_delete atomically update fd.tgid These functions need to set/reset the FD's tgid but when they're called there may still be wakeups on other threads that discover late updates and have to touch the tgid at the same time. As such, it is not possible to just read/write the tgid there. It must only be done using operations that are compatible with what other threads may be doing. As we're using inc/dec on the refcount, it's safe to AND the area to zero the lower part when resetting the value. However, in order to set the value, there's no other choice but fd_claim_tgid() which will assign it only if possible (via a CAS). This is convenient in the end because it protects the FD's masks from being modified by late threads, so while we hold this refcount we can safely reset the thread_mask and a few other elements. A debug test for non-null masks was added to fd_insert() as it must not be possible to face this situation thanks to the protection offered by the tgid.	2022-07-15 20:16:30 +02:00
Willy Tarreau	27a3245599	MEDIUM: fd: make fd_insert() take local thread masks fd_insert() was already given a thread group ID and a global thread mask. Now we're changing the few callers to take the group-local thread mask instead. It's passed directly into the FD's thread mask. Just like for previous commit, it must not change anything when a single group is configured.	2022-07-15 20:16:30 +02:00
Willy Tarreau	3638d174e5	MEDIUM: fd: make thread_mask now represent group-local IDs With the change that was started on other masks, the thread mask was still not fully converted, sometimes being used as a global mask and sometimes as a local one. This finishes the code modifications so that the mask is always considered as a group-local mask. This doesn't change anything as long as there's a single group, but is necessary for groups 2 and above since it's used against running_mask and so on.	2022-07-15 20:16:30 +02:00
Willy Tarreau	d6e1987612	MINOR: fd: make fd_clr_running() return the previous value instead It's an AND so it destroys information and due to this there's a call place where we have to perform two reads to know the previous value then to change it. With a fetch-and-and instead, in a single operation we can know if the bit was previously present, which is more efficient.	2022-07-15 20:16:30 +02:00
Willy Tarreau	a707d02657	MEDIUM: fd/poller: turn running_mask to group-local IDs From now on, the FD's running_mask only refers to local thread IDs. However, there remains a limitation, in updt_fd_polling(), we temporarily have to check and set shared FDs against .thread_mask, which still contains global ones. As such, nbtgroups > 1 may break (but this is not yet supported without special build options).	2022-07-15 20:16:30 +02:00
Willy Tarreau	6d3c501c08	MEDIUM: fd/poller: turn update_mask to group-local IDs From now on, the FD's update_mask only refers to local thread IDs. However, there remains a limitation, in updt_fd_polling(), we temporarily have to check and set shared FDs against .thread_mask, which still contains global ones. As such, nbtgroups > 1 may break (but this is not yet supported without special build options).	2022-07-15 20:16:30 +02:00
Willy Tarreau	63022128a5	MEDIUM: fd/poller: turn polled_mask to group-local IDs This changes the signification of each bit in the polled_mask so that now each bit represents a local thread ID for the current group instead of a global thread ID. As such, all tests now apply to ltid_bit instead of tid_bit. No particular check was made to verify that the FD's tgid matches the current one because there should be no case where this is not true. A check was added in epoll's __fd_clo() to confirm it never differs unless expected (soft stop under thread isolation, or master in starting mode going to exec mode), but that doesn't prevent from doing the job: it only consists in checking in the group's threads those that are still polling this FD and to remove them. Some atomic loads were added at the various locations, and most repetitive references to polled_mask[fd].xx were turned to a local copy instead making the code much more clear.	2022-07-15 20:16:30 +02:00
Willy Tarreau	0dc1cc93b6	MAJOR: fd: grab the tgid before manipulating running We now grab a reference to the FD's tgid before manipulating the running_mask so that we're certain it corresponds to our own group (hence bits), and we drop it once we've set the bit. For now there's no measurable performance impact in doing this, which is great. The lock can be observed by perf top as taking a small share of the time spent in fd_update_events(), itself taking no more than 0.28% of CPU under 8 threads. However due to the fact that the thread groups are not yet properly spread across the pollers and the thread masks are still wrong, this will trigger some BUG_ON() in fd_insert() after a few tens of thousands of connections when threads other than those of group 1 are reached, and this is expected.	2022-07-15 20:16:30 +02:00
Willy Tarreau	c243182370	MINOR: cli/fd: show fd's tgid and refcount in "show fd" We really need to display these values now.	2022-07-15 19:58:06 +02:00
Willy Tarreau	9464bb1f05	MEDIUM: fd: add the tgid to the fd and pass it to fd_insert() The file descriptors will need to know the thread group ID in addition to the mask. This extends fd_insert() to take the tgid, and will store it into the FD. In the FD, the tgid is stored as a combination of tgid on the lower 16 bits and a refcount on the higher 16 bits. This allows to know when it's really possible to trust the tgid and the running mask. If a refcount is higher than 1 it indeed indicates another thread else might be in the process of updating these values. Since a closed FD must necessarily have a zero refcount, a test was added to fd_insert() to make sure that it is the case.	2022-07-15 19:58:06 +02:00
Willy Tarreau	512dd2dc1c	MINOR: fd: make fd_insert() apply the thread mask itself It's a bit ugly to see that half of the callers of fd_insert() have to apply all_threads_mask themselves to the bit field they're passing, because usually it comes from a listener that may have other bits set. Let's make the function apply the mask itself.	2022-07-15 19:58:06 +02:00
Willy Tarreau	8e2c0fa8e5	MINOR: fd: delete unused updates on close() After a poller's ->clo() was called to completely terminate operations on an FD, there's no reason for keeping updates on this FD, so if any updates were already programmed it would be nice if we could delete them. Tests show that __fd_clo() is called roughly half of the time with the last FD from the local update list, which possibly makes sense if a close has to appear after a polling change resulting from an incomplete read or the end of a send(). We can detect this and remove the last entry, which gives less work to do during the update() call, and eliminates most of the poll_drop_fd event reports. Note that while tempting, this must not be backported because it's only safe to be done now that fd_delete_orphan() clears the update mask as we need to be certain not to miss it: - if the update mask is kept up with no entry, we can miss future updates ; - if the update mask is cleared too fast, it may result in failure to add a shared event.	2022-07-15 19:58:06 +02:00
Willy Tarreau	35ee710ece	MEDIUM: fd/poller: make the update-list per-group The update-list needs to be per-group because its inspection is based on a mask and we need to be certain when scanning it if a mask is for the same thread or another one. Once per-group there's no doubt about it, even if the FD's polling changes, the entry remains valid. It will be needed to check the tgid though. Note that a soft-stop or pause/resume might not necessarily work here with tgroups>1, because the operation might be delivered to a thread that doesn't belong to the group and whoe update mask will not reflect one that is interesting here. We can't do better at this stage.	2022-07-15 19:57:28 +02:00
Willy Tarreau	2f36d902aa	MAJOR: fd: remove pending updates upon real close Dealing with long-lasting updates that outlive a close() is always going to be quite a problem, not because of the thread that will discover such updates late, but mostly due to the shared update_list that will have an entry on hold making it difficult to reuse it, and requiring that the fd's tgid is changed and the update_mask reset from a safe location. After careful inspection, it turns out that all our pollers that support automatic event removal upon close() do not need any extra bookkeeping, and that poll and select that use an internal representation already provide a poller->clo() callback that is already used to update the local event. As such, it is already safe to reset the update mask and to remove the event from the shared list just before the final close, because nothing remains to be done with this FD by the poller. Doing so considerably simplifies the handling of updates, which will only have to be inspected by the pollers, while the writers can continue to consider that the entries are always valid. Another benefit is that it will be possible to reduce contention on the update_list by just having one update_list per group (left to be done later if needed).	2022-07-15 19:43:10 +02:00
Willy Tarreau	15c5500b6e	MEDIUM: conn: make conn_backend_get always scan the same group We don't want to pick idle connections from another thread group, this would be very slow by forcing to share undesirable data. This patch makes sure that we start seeking from the current thread group's threads only and loops over that range exclusively. It's worth noting that the next_takeover pointer remains per-server and will bounce when multiple groups use it at the same time. But we preserve the perturbation by applying a modulo when retrieving it, so that when groups are of the same size (most common case), the index will not even change. At this time it doesn't seem worth storing one index per group in servers, but that might be an option if any contention is detected later.	2022-07-15 19:43:10 +02:00
Willy Tarreau	91a7c164b4	MINOR: task: move the niced_tasks counter to the thread group context This one is only used as a hint to improve scheduling latency, so there is no more point in keeping it global since each thread group handles its own run q	2022-07-15 19:43:10 +02:00
Willy Tarreau	b0e7712fb2	MEDIUM: task/thread: move the task shared wait queues per thread group Their migration was postponed for convenience only but now's time for having the shared wait queues per thread group and not just per process, otherwise the WQ lock uses a huge amount of CPU alone.	2022-07-15 19:43:10 +02:00
Willy Tarreau	82e378aa8a	MINOR: fd/thread: get rid of thread_mask() Since commit `d2494e048` ("BUG/MEDIUM: peers/config: properly set the thread mask") there must not remain any single case of a receiver that is bound nowhere, so there's no need anymore for thread_mask(). We're adding a test in fd_insert() to make sure this doesn't happen by accident though, but the function was removed and its rare uses were replaced with the original value of the bind_thread msak.	2022-07-15 19:43:10 +02:00
Willy Tarreau	6bdf9452c0	MINOR: cli/threads: always bind CLI to thread group 1 When using multiple groups, the stats socket starts to emit errors and it's not natural to have to touch the global section just to specify "thread 1/all". Let's pre-attach these sockets to thread group 1. This will cause errors when trying to change the group but this really is not a problem for now as thread groups are not enabled by default. This will make sure configs remain portable and may possibly be relaxed later.	2022-07-15 19:43:10 +02:00
Willy Tarreau	dcbd763fe9	MINOR: mworker/threads: limit the mworker sockets to group 1 As a side effect of commit `34aae2fd1` ("MEDIUM: mworker: set the iocb of the socketpair without using fd_insert()"), a config may now refuse to start if there are multiple groups configured because the default bind mask may span over multiple groups, and it is not possible to force it to work differently. Let's just assign thread group 1 to the master<->worker sockets so that the thread bindings automatically resolve to a single group. The same was done for the master side of the socket even if it's not used. It will avoid being forgotten in the future.	2022-07-15 19:43:10 +02:00
Willy Tarreau	5b09341c02	MEDIUM: cpu-map: replace the process number with the thread group number The principle remains the same, but instead of having a single process and ignoring extra ones, now we set the affinity masks for the respective threads of all groups. The doc was updated with a few extra examples.	2022-07-15 19:43:10 +02:00
Willy Tarreau	1b2b59bfa7	MINOR: thread: remove MAX_THREADS limitation This one is now causing difficulties during the development phase and it's going to disappear anyway, let's get rid of it.	2022-07-15 19:43:10 +02:00
Willy Tarreau	e5715bface	MEDIUM: poller: disable thread-groups for poll() and select() These old legacy pollers are not designed for this. They're still using a shared list of events for all threads, this will not scale at all, so there's no point in enabling thread-groups there. Modern systems have epoll, kqueue or event ports and do not need these ones. We arrange for failing at boot time, only when thread-groups > 1 so that existing setups will remain unaffected. If there's a compelling reason for supporting thread groups with these pollers in the future, the rework should not be too hard, it would just consume a lot of memory to have an fd_evts[] array per thread, but that is doable.	2022-07-15 19:43:10 +02:00
Willy Tarreau	b1093c6ba2	MEDIUM: poller: program the update in fd_update_events() for a migrated FD When an FD is migrated, all pollers program an update. That's useless code duplication, and when thread groups will be supported, this will require an extra round of locking just to verify the update_mask on return. Let's just program the update direction from fd_update_events() as it already does for closed FDs, this becomes more logical.	2022-07-15 19:43:10 +02:00
Willy Tarreau	1b927eb3c3	MEDIUM: proto: stop protocols under thread isolation during soft stop protocol_stop_now() is called from do_soft_stop_now() running on any thread that received the signal. The problem is that it will call some listener handlers to close the FD, resulting in an fd_delete() being called from the wrong group. That's not clean and we cannot even rely on the thread mask to show up. One interesting long-term approach could be to have kill queues for FDs, and maybe we'll need them in the long run. However that doesn't work well for listeners in this situation. Let's simply isolate ourselves during this instant. We know we'll be alone dealing with the close and that the FD will be instantly deleted since not in use by any other thread. It's not the cleanest solution but it should last long enough without causing trouble.	2022-07-15 19:43:10 +02:00
Willy Tarreau	7aa41196cf	MEDIUM: debug/threads: make the lock debugging take tgroups into account Since we have to use masks to verify owners/waiters, we have no other option but to have them per group. This definitely inflates the size of the locks, but this is only used for extreme debugging anyway so that's not dramatic. Thus as of now, all masks in the lock stats are local bit masks, derived from ti->ltid_bit. Since at boot ltid_bit might not be set, we just take care of this situation (since some structs are initialized under look during boot), and use bit 0 from group 0 only.	2022-07-15 19:41:26 +02:00
Willy Tarreau	4d9888ca69	CLEANUP: fd: get rid of the __GET_{NEXT,PREV} macros They were initially made to deal with both the cache and the update list but there's no cache anymore and keeping them for the update list adds a lot of obfuscation that is really not desired. Let's get rid of them now. Their purpose was simply to get a pointer to fdtab[fd].update.{,next,prev} in order to perform atomic tests and modifications. The offset passed in argument to the functions (fd_add_to_fd_list() and fd_rm_from_fd_list()) was the offset of the ->update field in fdtab, and as it's not used anymore it was removed. This also removes a number of casts, though those used by the atomic ops have to remain since only scalars are supported.	2022-07-15 19:41:26 +02:00
Willy Tarreau	740038c8b9	MINOR: listener/config: make "thread" always support up to LONGBITS The difference is subtle but in one place there was MAXTHREADS and this will not work anymore once it goes over 64.	2022-07-15 19:41:26 +02:00
Willy Tarreau	acd644197f	MEDIUM: config: remove the "process" keyword on "bind" lines It was deprecated, marked for removal in 2.7 and was already emitting a warning, let's get rid of it. Note that we've kept the keyword detection to suggest to use "thread" instead.	2022-07-15 19:41:26 +02:00
Willy Tarreau	94f763b5e4	MEDIUM: config: remove deprecated "bind-process" directives from frontends This was already causing a deprecation warning and was marked for removal in 2.7, now it happens. An error message indicates this doesn't exist anymore.	2022-07-15 19:41:26 +02:00
Willy Tarreau	91f7a1af34	CLEANUP: applet: remove the obsolete command context from the appctx The "ctx" and "st2" parts in the appctx were marked for removal in 2.7 and were emulated using memcpy/memset etc for possible external code. Let's remove this now.	2022-07-15 19:41:26 +02:00
Willy Tarreau	9a7fa90239	MINOR: cli/activity: add a thread number argument to "show activity" The output of "show activity" can be so large that the output is visually unreadable on a screen. Let's add an option to filter on the desired column (actually the thread number), use "0" to report only the first column (aggregated/sum/avg), and use "-1", the default, for the normal detailed dump.	2022-07-15 19:41:26 +02:00
Willy Tarreau	dadf00e226	DEBUG: cli: add a new "debug dev deadlock" expert command This command will create the requested number of tasks competing on a lock, resulting in triggering the watchdog and crashing the process. This will help stress the watchdog and inspect the lock debugging parts.	2022-07-15 19:41:26 +02:00
Willy Tarreau	dd75b64cdf	MINOR: cli/streams: show a stream's tgid next to its thread ID We now display both the global thread ID and the tgid/ltid pair so that it's easier to match it with the FD.	2022-07-15 19:41:26 +02:00
Willy Tarreau	f0c86ddfe8	BUG/MEDIUM: debug: fix parallel thread dumps again The previous attempt to fix thread dumps in commit `672972604` ("BUG/MEDIUM: debug: fix possible hang when multiple threads dump at once") still had some shortcomings. Sometimes parallel dumps are jerky essentially due to the way that threads synchronize on startup and end. In addition the risk of waiting forever for a stopped thread exists, and panics happening in parallel to thread dumps are not more reliable either. This commit revisits the state transitions so that all threads may request a dump in parallel, that all of them wait for each other in the handler, and that one thread is responsible for counting every other and checking that the total matches the number of active threads. Then for stopping there's a finishing phase that all threads wait for so that none quits this area too early. Given that we now know the number of participants to the dump, we can let them each decrement the counter when leaving so that another dump may only start after the last participant has completely left. Now many thread dumps in parallel are running fine, so do panics. No backport is needed as this was the result of the changes for thread groups.	2022-07-15 19:41:26 +02:00
Willy Tarreau	55433f9b34	BUG/MINOR: debug: enter ha_panic() only once Some panic dumps are mangled or truncated due to the watchdog firing at the same time on multiple threads and calling ha_panic() simultaneously. What may happen in this case is that the second one waits for the first one to finish but as soon as it's done the second one resets the buffer and dumps again, sometimes resetting the first one's dump. Also the first one's abort() may trigger while the second one is currently dumping, resulting in a full dump followed by a truncated one, leading to confusion. Sometimes some lines appear in the middle of a dump as well. It doesn't happen often and is easier to trigger by causing massive deadlocks. There's no reason for the process to resist to a panic, so we can safely add a counter and no nothing on subsequent calls. Ideally we'd wait there forever but as this may happen inside a signal handler (e.g. watchdog), it doesn't always work, so the easiest thing to do is to return so that the thread is interrupted as soon as possible and brought to the debug handler to be dumped. This should be backported, at least to 2.6 and possibly to older versions as well.	2022-07-15 19:41:26 +02:00
Willy Tarreau	f15c75a2d3	BUG/MINOR: thread: use the correct thread's group in ha_tkillall() In ha_tkillall(), the current thread's group was used to check for the thread being running instead of using the target thread's group mask. Most of the time it would not have any effect unless some groups are uneven where it can lead to incomplete thread dumps for example. No backport is needed, this is purely 2.7.	2022-07-15 19:41:26 +02:00
Willy Tarreau	52f238d326	BUG/MEDIUM: cli/threads: make "show threads" more robust on applets Running several concurrent "show threads" in loops might occasionally cause a segfault when trying to retrieve the stream from appctx_sc() which may be null while the applet is finishing. It's not easy to reproduce, it requires 3-5 sessions in parallel for about a minute or so. The appctx_sc must be checked before passing it to sc_strm(). This must be backported to 2.6 which also has the bug.	2022-07-15 19:41:26 +02:00
Willy Tarreau	9b0f0d146f	BUG/MINOR: threads: produce correct global mask for tgroup > 1 In thread_resolve_group_mask(), if a global thread number is passed and it belongs to a group greater than 1, an incorrect shift resulted in shifting that ID again which made it appear nowhere or in a wrong group possibly. The bug was introduced in 2.5 with commit `627def9e5` ("MINOR: threads: add a new function to resolve config groups and masks") though the groups only starts to be usable in 2.7, so there is no impact for this bug, hence no backport is needed.	2022-07-15 19:41:26 +02:00
Amaury Denoyelle	114c9c87ce	MINOR: h3: implement graceful shutdown with GOAWAY Implement graceful shutdown as specified in RFC 9114. A GOAWAY frame is generated with stream ID to indicate range of processed requests. This process is done via the release app protocol operation. The MUX is responsible to emit the generated GOAWAY frame after app release. A CONNECTION_CLOSE will be emitted once there is no unacknowledged STREAM frames.	2022-07-15 15:56:13 +02:00
Amaury Denoyelle	d701039773	MINOR: h3: store control stream in h3c Store a reference to the HTTP/3 control stream in h3c context. This will be useful to implement GOAWAY emission without having to store the control stream ID on opening.	2022-07-15 15:56:13 +02:00
Amaury Denoyelle	a154dc0290	MINOR: mux-quic: send one last time before release Call qc_send() on qc_release(). This is mostly useful for application protocol with a connection closing procedure. Most notably, this will be useful to implement HTTP/3 GOAWAY emission.	2022-07-15 15:56:13 +02:00
Amaury Denoyelle	c49d5d1a4b	CLEANUP: mux-quic: move qc_release() This change is purely cosmetic. qc_release() function is moved just before qc_io_cb(). It's cleaner as it brings it closer where it is used. More importantly, this will be required to be able to use it in qc_send() function.	2022-07-15 15:56:13 +02:00
Amaury Denoyelle	240b1b108b	MEDIUM: quic: send CONNECTION_CLOSE on released MUX Send a CONNECTION_CLOSE if the MUX has been released and all STREAM data are acknowledged. This is useful to prevent a client from trying to use a connection which have the upper layer closed. To implement this a new function qc_check_close_on_released_mux() has been added. It is called on QUIC MUX release notification and each time a qc_stream_desc has been released. This commit is associated with the previous one : MINOR: mux-quic/h3: schedule CONNECTION_CLOSE on app release Both patches are required to prevent the risk of browsers stuck on webpage loading if MUX has been released. On CONNECTION_CLOSE reception, the client will reopen a new QUIC connection.	2022-07-15 15:56:13 +02:00
Amaury Denoyelle	069288b4c0	MINOR: mux-quic/h3: prepare CONNECTION_CLOSE on release When MUX is released, a CONNECTION_CLOSE frame should be emitted. This will ensure that the client does not use anymore a half-dead connection. App protocol layer is responsible to provide the error code via release callback. For HTTP/3 NO_ERROR is used as specified in RFC 9114. If no release callback is provided, generic QUIC NO_ERROR code is used. Note that a graceful shutdown is used : quic_conn must emit CONNECTION_CLOSE frame when possible. This will be provided in another patch. This change should limit the risk of browsers stuck on webpage loading if MUX has been released. On CONNECTION_CLOSE reception, the client will reopen a new QUIC connection.	2022-07-15 15:20:33 +02:00
Amaury Denoyelle	d666d740d2	MINOR: mux-quic: support app graceful shutdown Adjust qcc_emit_cc_app() to allow the delay of emission of a CONNECTION_CLOSE. This will only set the error code but the quic-conn layer is not flagged for immediate close. The quic-conn will be responsible to shut the connection when deemed suitable. This change will allow to implement application graceful shutdown, such as HTTP/3 with GOAWAY emission. This will allow to emit closing frames on MUX release. Once all work is done at the lower layer, the quic-conn should emit a CONNECTION_CLOSE with the registered error code.	2022-07-15 15:06:59 +02:00
Amaury Denoyelle	57e6db7021	MINOR: quic: define a generic QUIC error type Define a new structure quic_err to abstract a QUIC error type. This allows to easily differentiate a transport and an application error code. This simplifies error transmission from QUIC MUX and H3 layers. This new type is defined in quic_frame module. It is used to replace <err_code> field in <quic_conn>. QUIC_FL_CONN_APP_ALERT flag is removed as it is now useless. Utility functions are defined to be able to quickly instantiate transport, tls and application errors.	2022-07-15 14:57:49 +02:00
Amaury Denoyelle	72d86509f1	BUG/MINOR: quic: fix closing state on NO_ERROR code sent Reception is disabled as soon as a CONNECTION_CLOSE emission is required. An early return is done on qc_lstnr_pkt_rcv() to implement this. This condition is not functional if the error code sent is NO_ERROR (0x00). To fix this, check the quic-conn flags instead of the error code. Currently this bug has no impact has NO_ERROR emission is not used. This can be backported up to 2.6.	2022-07-13 15:33:15 +02:00
Willy Tarreau	672972604f	BUG/MEDIUM: debug: fix possible hang when multiple threads dump at once A bug in the thread dumper was introduced by commit `00c27b50c` ("MEDIUM: debug: make the thread dumper not rely on a thread mask anymore"). If two or more threads try to trigger a thread dump exactly at the same time, the second one may loop indefinitely trying to set the value to 1 while the other ones will wait for it to finish dumping before leaving. This is a consequence of a logic change using thread numbers instead of a thread mask, as threads do not need to see all other ones there anymore. No backport is needed, this is only for 2.7.	2022-07-13 09:03:02 +02:00
Amaury Denoyelle	a5b5075211	MEDIUM: mux-quic: implement STOP_SENDING handling Implement support for STOP_SENDING frame parsing. The stream is resetted as specified by RFC 9000. This will automatically interrupt all future send operation in qc_send(). A RESET_STREAM will be sent with the code extracted from the original STOP_SENDING frame.	2022-07-11 16:45:04 +02:00
Amaury Denoyelle	843a1196b3	MEDIUM: mux-quic: implement RESET_STREAM emission Implement functions to be able to reset a stream via RESET_STREAM. If needed, a qcs instance is flagged with QC_SF_TO_RESET to schedule a stream reset. This will interrupt all future send operations. On stream emission, if a stream is flagged with QC_SF_TO_RESET, a RESET_STREAM frame is generated and emitted to the transport layer. If this operation succeeds, the stream is locally closed. If upper layer is instantiated, error flag is set on it.	2022-07-11 16:45:04 +02:00
Amaury Denoyelle	20d1f84ce4	MINOR: mux-quic: use stream states to mark as detached Adjust condition to detach a qcs instance : if the stream is not locally close it is not directly free. This should improve stream closing by ensuring that either FIN or a RESET_STREAM is sent before destroying it.	2022-07-11 16:41:10 +02:00
Amaury Denoyelle	38e6006da1	MINOR: mux-quic: define basic stream states Implement a basic state machine to represent stream lifecycle. By default a stream is idle. It is marked as open when sending or receiving the first data on a stream. Bidirectional streams has two states to represent the closing on both receive and send channels. This distinction does not exists for unidirectional streams which passed automatically from open to close state. This patch is mostly internal and has a limited visible impact. Some behaviors are slightly updated : * closed streams are garbage collected at the start of io handler * send operation is interrupted if a stream is close locally Outside of this, there is no functional change. However, some additional BUG_ON guards are implemented to ensure that we do not conduct invalid operation on a stream. This should strengthen the code safety. Also, stream states are displayed on trace which should help debugging.	2022-07-11 16:37:21 +02:00
Amaury Denoyelle	b68559a9aa	MINOR: mux-quic: support stream opening via MAX_STREAM_DATA MAX_STREAM_DATA can be used as the first frame of a stream. In this case, the stream should be opened, if it respects flow-control limit. To implement this, simply replace plain lookup in stream tree by qcc_get_qcs() at the start of the parsing function. This automatically takes care of opening the stream if not already done. As specified by RFC 9000, if MAX_STREAM_DATA is receive for a receive-only stream, a STREAM_STATE_ERROR connection error is emitted.	2022-07-11 16:24:03 +02:00
Amaury Denoyelle	57161b7d0c	MINOR: mux-quic: do not ack STREAM frames on unrecoverable error Improve return path for qcc_recv() on STREAM parsing. It returns 0 on success. On error, a non-zero value is returned which indicates to the caller that the packet containing the frame should not be acknowledged. When qcc_recv() generates a CONNECTION_CLOSE or RESET_STREAM, either directly or via qcc_get_qcs(), an error is returned which ensure that no acknowledgement is generated. This required an adjustment on qcc_get_qcs() API which now returns a success/error code. The stream instance is returned via a new out argument.	2022-07-11 16:24:03 +02:00
Amaury Denoyelle	5fbb8691d4	MINOR: mux-quic: filter send/receive-only streams on frame parsing Extend the function qcc_get_qcs() to be able to filter send/receive-only unidirectional streams. A connection error STREAM_STATE_ERROR is emitted if this new filter does not match. This will be useful when various frames handlers are converted with qcc_get_qcs(). Depending on the frame type, it will be easy to filter on the forbidden stream types as specified in RFC 9000.	2022-07-11 16:24:03 +02:00
Amaury Denoyelle	4561f84ad4	MINOR: mux-quic: implement qcs_alert() Implement a simple function to notify a possible subscriber or wake up the upper layer if a special condition happens on a stream. For the moment, this is only used to replace identical code in qc_wake_some_streams().	2022-07-11 16:24:03 +02:00
Amaury Denoyelle	392e94e985	MINOR: mux-quic: add traces on frame parsing functions Add traces for parsing functions for MAX_DATA and MAX_STREAM_DATA.	2022-07-11 16:24:03 +02:00
Amaury Denoyelle	c1a6dfd477	MINOR: mux-quic: rename stream purge function Rename qc_release_detached_streams() to qc_purge_streams(). The aim is to have a more generic name. It's expected to complete this function to add other criteria to purge dead streams. Also the function documentation has been corrected. It does not return a number of streams. Instead it is a boolean value, to true if at least one stream was released.	2022-07-11 16:24:03 +02:00
Amaury Denoyelle	b143723411	REORG: mux-quic: rename stream initialization function Rename both qcc_open_stream_local/remote() functions to qcc_init_stream_local/remote(). This change is purely cosmetic. It will reduces the ambiguity with the soon to be implemented OPEN states for QCS instances.	2022-07-11 16:24:03 +02:00
Amaury Denoyelle	e53b489826	BUG/MEDIUM: mux-quic: fix server chunked encoding response QUIC MUX was not able to correctly deal with server response using chunked transfer-encoding. All data will be transfered correctly to the client but the FIN bit is missing. The transfer will never stop as the client will wait indefinitely for the FIN bit. This bug happened because the HTX message representing a chunked encoded payload contains a final empty block with the EOM flag. However, emission is skipped by QUIC MUX if there is no data to transfer. To fix this, the condition was completed to ensure that there is no need to send the FIN signal. If this is false, data emission will proceed even if there is no data : this will generate an empty QUIC STREAM frame with FIN set which will mark the end of the transfer. To ensure that a FIN STREAM frame is sent only one time, QC_SF_FIN_STREAM is resetted on send confirmation from the transport in qcc_streams_sent_done(). This bug was reproduced when dealing with chunked transfer-encoding response for the HTTP server. This must be backported up to 2.6.	2022-07-11 16:21:52 +02:00
Willy Tarreau	a88e8bf428	BUILD: http: silence an uninitialized warning affecting gcc-5 When building with gcc-5, one can see this warning: src/http_fetch.c: In function 'smp_fetch_meth': src/http_fetch.c:356:6: warning: 'htx' may be used uninitialized in this function [-Wmaybe-uninitialized] sl = http_get_stline(htx); ^ It's wrong since the only way to reach this code is to have met the same condition a few lines before and initialized the htx variable. The reason in fact is that the same test happens on different variables of distinct types, so the compiler possibly doesn't know that the condition is the same. Newer gcc versions do not have this problem. Let's just move the assignment earlier and have the exact same test, as it's sufficient to shut this up. This may have to be backported to 2.6 since the code is the same there.	2022-07-10 14:13:48 +02:00
Willy Tarreau	0d023774bf	MEDIUM: epoll: don't synchronously delete migrated FDs Between 1.8 and 1.9 commit `d9e7e36c6` ("BUG/MEDIUM: epoll/threads: use one epoll_fd per thread") split the epoll poller to use one poller per thread (and this was backported to 1.8). This patch added a call to epoll_ctl(DEL) on return from the I/O handler as a safe way to deal with a detected thread migration when that code was still quite fragile. One aspect of this choice was that by then we wanted to maintain support for the rare old bogus epoll implementations that failed to remove events on close(), so risking to lose the event was not an option. Later in 2.5, commit `200bd50b7` ("MEDIUM: fd: rely more on fd_update_events() to detect changes") changed the code to perform most of the operations inside fd_update_events(), but it maintained that oddity, to the point that strictly all pollers except epoll now just add an update to be dealt with at the next round. This approach is much more efficient, because under load and server-side connection reuse, it's perfectly possible for a thread to see the same FD several times in a poll loop, the first time to relinquish it after a migration, then the other thread makes a request, gets its response, and still during the same loop for the first one, grabbing an idle connection to send a request and wait for a response will program a new update on this FD. By using a synchronous epoll_ctl(DEL), we effectively lose the opportunity to aggregate certain changes in the same update. Some tests performed locally with 8 threads and one server show that on average, by using an update instead of a synchronous call, we reduce the number of epoll_ctl() calls by 25-30% (under low loads it will probably not change anything). So this patch implements the same method for all pollers and replaces the synchronous epoll_ctl() with an update.	2022-07-10 14:13:48 +02:00
Christopher Faulet	372b38f935	BUG/MEDIUM: mux-h1: Handle connection error after a synchronous send Since commit `d1480cc8` ("BUG/MEDIUM: stream-int: do not rely on the connection error once established"), connection errors are not handled anymore by the stream-connector once established. But it is a problem for the H1 mux when an error occurred during a synchronous send in h1_snd_buf(). Because in this case, the connction error is just missed. It leads to a session leak until a timeout is reached (client or server). To fix the bug, the connection flags are now checked in h1_snd_buf(). If there is an error, it is reported to the stconn level by setting SF_FL_ERROR flags. But only if there is no pending data in the input buffer. This patch should solve the issue #1774. It must be backported as far as 2.2.	2022-07-08 16:37:31 +02:00
Christopher Faulet	52fc0cbaad	BUG/MEDIUM: http-ana: Don't wait to have an empty buf to switch in TUNNEL state When we want to establish a tunnel on a side, we wait to have flush all data from the buffer. At this stage the other side is at least in DONE state. But there is no real reason to wait. We are already in DONE state on its side. So all the HTTP message was already forwarded or planned to be forwarded. Depending on the scheduling if the mux already started to transfer tunneled data, these data may block the switch in TUNNEL state and thus block these data infinitly. This bug exists since the early days of HTX. May it was mandatory but today it seems useless. But I honestly don't remember why this prerequisite was added. So be careful during the backports. This patch should be backported with caution. At least as far as 2.4. For 2.2 and 2.0, it seems to be mandatory too. But a review must be performed.	2022-07-08 16:37:31 +02:00
Christopher Faulet	5966e40641	BUG/MINOR: mux-h1: Be sure to commit htx changes in the demux buffer When a buffer area is casted to an htx message, depending on the method used, the underlying buffer may be updated or not. The htxbuf() function does not change the buffer state. We assume the buffer was already prepared to store an htx message. htx_from_buf() on its side, updates the buffer. With the first function, we only need to commit changes to the underlying buffer if the htx message is changed. With last one, we must always commit the changes. The idea is to be sure to keep non-empty HTX messages while an empty message must be lead to an empty buffer after commit. All that said because in h1_process_demux(), the changes is not always committed as expected. When the demux is blocked, we just leave the function. So it is possible to have an empty htx message stored in a buffer that appears non-empty. It is not critical, but the buffer cannot be released in this state. And we should always release an empty buffer. This patch must be backported as far as 2.4.	2022-07-08 16:37:31 +02:00
William Lallemand	a46a99e98c	MEDIUM: mworker/systemd: send STATUS over sd_notify The sd_notify API is not able to change the "Active:" line in "systemcl status". However a message can still be displayed on a "Status: " line, even if the service is still green and "active (running)". When startup succeed the Status will be set to "Ready.", upon a reload it will be set to "Reloading Configuration." If the configuration succeed "Ready." again. However if the reload failed, it will be set to "Reload failed!". Keep in mind that the "Active:" line won't change upon a reload failure, and will still be green.	2022-07-07 14:48:46 +02:00
Christopher Faulet	12f6dbb863	BUG/MEDIUM: http-fetch: Don't fetch the method if there is no stream The "method" sample fetch does not perform any check on the stream existence before using it. However, for errors triggered at the mux level, there is no stream. When the log message is formatted, this sample fetch must fail. It must also fail when it is called from a health-check. This patch must be backported as far as 2.4.	2022-07-07 09:35:58 +02:00
Christopher Faulet	d1d983fb12	MINOR: http-htx: Use new HTTP functions for the scheme based normalization Use http_get_host_port() and http_is_default_port() functions to perform the scheme based normalization.	2022-07-07 09:35:58 +02:00
Christopher Faulet	3f5fbe9407	BUG/MEDIUM: h1: Improve authority validation for CONNCET request From time to time, users complain to get 400-Bad-request responses for totally valid CONNECT requests. After analysis, it is due to the H1 parser performs an exact match between the authority and the host header value. For non-CONNECT requests, it is valid. But for CONNECT requests the authority must contain a port while it is often omitted from the host header value (for default ports). So, to be sure to not reject valid CONNECT requests, a basic authority validation is now performed during the message parsing. In addition, the host header value is normalized. It means the default port is removed if possible. This patch should solve the issue #1761. It must be backported to 2.6 and probably as far as 2.4.	2022-07-07 09:35:58 +02:00
Christopher Faulet	ca7218aaf0	MINOR: http: Add function to detect default port http_is_default_port() can be used to test if a port is a default HTTP/HTTPS port. A scheme may be specified. In this case, it is used to detect defaults ports, 80 for "http://" and 443 for "https://". Otherwise, with no scheme, both are considered as default ports.	2022-07-06 17:54:03 +02:00
Christopher Faulet	658f971621	MINOR: http: Add function to get port part of a host http_get_host_port() function can be used to get the port part of a host. It will be used to get the port of an uri authority or a host header value. This function only look for a port starting from the end of the host. It is the caller responsibility to call it with a valid host value. An indirect string is returned.	2022-07-06 17:54:03 +02:00
Christopher Faulet	0eab050b04	BUG/MINOR: http-htx: Fix scheme based normalization for URIs wih userinfo The scheme based normalization is not properly handled the URI's userinfo, if any. First, the authority parser is not called with "no_userinfo" parameter set. Then it is skipped from the URI normalization. This patch must be backported as far as 2.4.	2022-07-06 17:54:02 +02:00
William Lallemand	1d93217a05	BUG/MINOR: peers: fix possible NULL dereferences at config parsing Patch `49f6f4b` ("BUG/MEDIUM: peers: fix segfault using multiple bind on peers sections") introduced possible NULL dereferences when parsing the peers configuration. Fix the issue by checking the return value of bind_conf_uniq_alloc(). This patch should be backported as far as 2.0.	2022-07-06 14:40:11 +02:00
Willy Tarreau	ad92fdf196	CLEANUP: thread: also remove a thread's bit from stopping_threads on stop As much as possible we should take care of not leaving bits from stopped threads in shared thread masks. It can avoid issues like the previous fix and will also make debugging less confusing.	2022-07-06 10:19:46 +02:00
Willy Tarreau	f34a3fa33d	BUG/MEDIUM: thread: mask stopping_threads with threads_enabled when checking it When soft-stopping, there's a comparison between stopping_threads and threads_enabled to make sure all threads are stopped, but this is not correct and is racy since the threads_enabled bit is removed when a thread is stopped but not its stopping_threads bit. The consequence is that depending on timing, when stopping, if the first stopping thread is fast enough to remove its bit from threads_enabled, the other threads will see that stopping_threads doesn't match threads_enabled anymore and will wait forever. As such the mask must be applied to stopping_threads during the test. This issue was introduced in recent commit `ef422ced9` ("MEDIUM: thread: make stopping_threads per-group and add stopping_tgroups"), no backport is needed.	2022-07-06 10:19:46 +02:00
Christopher Faulet	4c3d3d2a68	BUG/MINOR: http-act: Properly generate 103 responses when several rules are used When several "early-hint" rules are used, we try, as far as possible, to merge links into the same 103-early-hints response. However, it only works if there is no ACLs. If a "early-hint" rule is not executed an invalid response is generated. the EOH block or the start-line may be missing, depending on the rule order. To fix the bug, we use the transaction status code. It is unused at this stage. Thus, it is set to 103 when a 103-early-hints response is in progress. And it is reset when the response is forwarded. In addition, the response is forwarded if the next rule is an "early-hint" rule with an ACL. This way, the response is always valid. This patch must be backported as far as 2.2.	2022-07-06 09:37:43 +02:00
Christopher Faulet	4c8e58def6	BUG/MINOR: http-check: Preserve headers if not redefined by an implicit rule When an explicit "http-check send" rule is used, if it is the first one, it is merge with the implicit rule created by "option httpchk" statement. The opposite is also true. Idea is to have only one send rule with the merged info. It means info defined in the second rule override those defined in the first one. However, if an element is not defined in the second rule, it must be ignored, keeping this way info from the first rule. It works as expected for the method, the uri and the request version. But it is not true for the header list. For instance, with the following statements, a x-forwarded-proto header is added to healthcheck requests: option httpchk http-check send meth GET hdr x-forwarded-proto https while by inverting the statements, no extra headers are added: http-check send meth GET hdr x-forwarded-proto https option httpchk Now the old header list is overriden if the new one is not empty. This patch should fix the issue #1772. It must be backported as far as 2.2.	2022-07-06 09:35:13 +02:00
Christopher Faulet	f0196f4f71	CLEANUP: bwlim: Set pointers to NULL when memory is released Calls to free() are replaced by ha_free(). And otherwise, the pointers are explicitly set to NULL after a release. There is no issue here but it could help debugging sessions.	2022-07-06 09:34:54 +02:00
Willy Tarreau	d2494e0489	BUG/MEDIUM: peers/config: properly set the thread mask The peers didn't have their bind_conf thread mask nor group set, because they're still not part of the global proxy list. Till 2.6 it seems it does not have any visible impact, since most listener-oriented operations pass through thread_mask() which detects null masks and turns them to all_threads_mask. But starting with 2.7 it becomes a problem as won't permit these null masks anymore. This patch duplicates (yes, sorry) the loop that applies to the frontend's bind_conf, though it is simplified (no sharding, etc). As the code is right now, it simply seems impossible to trigger the second (and largest) part of the check when leaving thread_resolve_group_mask() on success, so it looks like it might be removed. No backport is needed, unless a report in 2.6 or earlier mentions an issue with a null thread_mask.	2022-07-05 19:10:26 +02:00
Willy Tarreau	8d158132bd	BUG/MINOR: peers/config: always fill the bind_conf's argument Some generic frontend errors mention the bind_conf by its name as "bind '%s'", but if this is used on peers "bind" lines it shows "(null)" because the argument is set to NULL in the call to bind_conf_uniq_alloc() instead of passing the argument. Fortunately that's trivial to fix. This may be backported to older versions.	2022-07-05 19:06:47 +02:00
Amaury Denoyelle	bf91e3922b	MINOR: mux-quic: emit FINAL_SIZE_ERROR on invalid STREAM size Add a check on stream size when the stream is in state Size Known. In this case, a STREAM frame cannot change the stream size. If this is not respected, a CONNECTION_CLOSE with FINAL_SIZE_ERROR will be emitted as specified in the RFC 9000.	2022-07-05 16:44:01 +02:00
Amaury Denoyelle	3f39b40fe0	MINOR: mux-quic: rename qcs flag FIN_RECV to SIZE_KNOWN Rename QC_SF_FIN_RECV to the more generic name QC_SF_SIZE_KNOWN. This better align with the QUIC RFC 9000 which uses the "Size Known" state definition. This change is purely cosmetic.	2022-07-05 16:18:27 +02:00
Amaury Denoyelle	a509ffb505	MEDIUM: mux-quic: refactor streams opening Review the whole API used to access/instantiate qcs. A public function qcc_open_stream_local() is available to the application protocol layer. It allows to easily opening a local stream. The ID is automatically attributed to the next one available. For remote streams, qcc_open_stream_remote() has been implemented. It will automatically take care of allocating streams in a linear way according to the ID. This function is called via qcc_get_qcs() which can be used for each qcc_recv*() operations. For the moment, it is only used for STREAM frames via qcc_recv(), but soon it will be implemented for other frames types which can also be used to open a new stream. qcs_new() and qcs_free() has been restricted to the MUX QUIC only as they are now reserved for internal usage. This change is a pure refactoring and should not have any noticeable impact. It clarifies the developer intent and help to ensure that a stream is not automatically opened when not desired.	2022-07-05 16:18:27 +02:00
Amaury Denoyelle	3abeb57909	MINOR: mux-quic: implement accessor for sedesc Implement a function <qcs_sc> to easily access to the stconn associated with a QCS. This takes care of qcs.sd which may be NULL, for example for unidirectional streams. It is expected that in the future when implementing STOP_SENDING/RESET_STREAM, stconn must be notify about the event. This accessor will allow to easily test if the stconn is instantiated or not.	2022-07-05 11:22:15 +02:00
Amaury Denoyelle	a441ec9c7a	CLEANUP: mux-quic: do not export qc_get_ncbuf qc_get_ncbuf() is only used internally : thus its prototype in QUIC MUX include is not required.	2022-07-05 11:06:52 +02:00
William Lallemand	2ee490f613	CLEANUP: mworker: rename mworker_pipe to mworker_sockpair The function mworker_pipe_register_per_thread() is called this way because the master first used pipes instead of socketpairs. Rename mworker_pipe_register_per_thread() to mworker_sockpair_register_per_thread() in order to be more consistent. Also update a comment inside the function.	2022-07-05 09:06:04 +02:00
William Lallemand	34aae2fd12	MEDIUM: mworker: set the iocb of the socketpair without using fd_insert() The worker was previously changing the iocb of the socketpair in the worker by mworker_accept_wrapper(). However, it was done using fd_insert() instead of changing directly the callback in the fdtab[].iocb pointer. This patch cleans up this by part by removing fd_insert(). It also stops setting tid_bit on the thread mask, the socketpair will be handled by any thread from now.	2022-07-05 05:18:51 +02:00
Willy Tarreau	7509ec369a	MINOR: proxy: use tg->threads_enabled in hard_stop() to detect stopped threads Let's rely on tg->threads_enabled there to detect running threads. We should probably have a dedicated function for this in order to simplify the code and avoid the risk of using the wrong group ID.	2022-07-04 14:09:39 +02:00
Willy Tarreau	24cfc9f76e	BUG/MEDIUM: thread: check stopping thread against local bit and not global one Commit `ef422ced9` ("MEDIUM: thread: make stopping_threads per-group and add stopping_tgroups") moved the stopping_threads mask to per-group, but one test in the loop preserved its global value instead, resulting in stopping threads never sleeping on stop and eating 100% CPU until all were stopped. No backport is needed.	2022-07-04 14:09:39 +02:00
Willy Tarreau	291f6ff885	BUG/MEDIUM: threads: fix incorrect thread group being used on soft-stop Commit `377e37a80` ("MINOR: tinfo: add the mask of enabled threads in each group") forgot -1 on the tgid, thus the groups was not always correctly tested, which is visible only when running with more than one group. No backport is needed.	2022-07-04 13:37:31 +02:00
Willy Tarreau	89ed89e895	BUILD: debug: re-export thread_dump_state Building with threads and without thread dump (e.g. macos, freebsd) warns that thread_dump_state is unused. This happened in fact with recentcommit `1229ef312` ("MINOR: wdt: do not rely on threads_to_dump anymore"). The solution would be to mark it unused, but after a second thought, it can be convenient to keep it exported to help debug crashes, so let's export it again. It's just not referenced in include files since it's not needed outside.	2022-07-01 21:18:03 +02:00
Willy Tarreau	039972b4e5	BUILD: debug: fix build issue on clang with previous commit Since the thread_dump_state type changed to uint, the old value in the CAS needs to be the same as well.	2022-07-01 19:37:42 +02:00
Willy Tarreau	00c27b50c0	MEDIUM: debug: make the thread dumper not rely on a thread mask anymore The thread mask is too short to dump more than 64 bits. Thus here we're using a different approach with two counters, one for the next thread ID to dump (which always exists, as it's looked up), and the second one for the number of threads done dumping. This allows to dump threads in ascending order then to let them wait for all others to be done, then to leave without the risk of an overlapping dump until the done count is null again. This allows to remove threads_to_dump which was the last non-FD variable using a global thread mask.	2022-07-01 19:31:39 +02:00
Willy Tarreau	1229ef312d	MINOR: wdt: do not rely on threads_to_dump anymore This flag is not needed anymore as we're already marking the waiting threads as harmless, thus the thread's bit is already covered by this information. The variable was unexported.	2022-07-01 19:26:35 +02:00
Willy Tarreau	f7afdd910b	MINOR: debug: mark oneself harmless while waiting for threads to finish The debug_handler() function waits for other threads to join, but does not mark itself as harmless, so if at the same time another thread tries to isolate, this may deadlock. In practice this does not happen as the signal is received during epoll_wait() hence under harmless mode, but it can possibly arrive under other conditions. In order to improve this, while waiting for other threads to join, we're now marking the current thread as harmless, as it's doing nothing but waiting for the other ones. This way another harmless waiter will be able to proceed. It's valid to do this since we're not doing anything else in this loop. One improvement could be to also check for the thread being idle and marking it idle in addition to harmless, so that it can even release a full isolation requester. But that really doesn't look worth it.	2022-07-01 19:26:35 +02:00
Willy Tarreau	a2b8ed4b44	MINOR: thread: add is_thread_harmless() to know if a thread already is harmless The harmless status is not re-entrant, so sometimes for signal handling it can be useful to know if we're already harmless or not. Let's add a function doing that, and make the debugger use it instead of manipulating the harmless mask.	2022-07-01 19:26:35 +02:00
Willy Tarreau	598cf3f22e	MAJOR: threads: change thread_isolate to support inter-group synchronization thread_isolate() and thread_isolate_full() were relying on a set of thread masks for all threads in different states (rdv, harmless, idle). This cannot work anymore when the number of threads increases beyond LONGBITS so we need to change the mechanism. What is done here is to have a counter of requesters and the number of the current isolated thread. Threads which want to isolate themselves increment the request counter and wait for all threads to be marked harmless (or idle) by scanning all groups and watching the respective masks. This is possible because threads cannot escape once they discover this counter, unless they also want to isolate and possibly pass first. Once all threads are harmless, the requesting thread tries to self-assign the isolated thread number, and if it fails it loops back to checking all threads. If it wins it's guaranted to be alone, and can drop its harmless bit, so that other competing threads go back to the loop waiting for all threads to be harmless. The benefit of proceeding this way is that there's very little write contention on the thread number (none during work), hence no cache line moves between caches, thus frozen threads do not slow down the isolated one. Once it's done, the isolated thread resets the thread number (hence lets another thread take the place) and decrements the requester count, thus possibly releasing all harmless threads. With this change there's no more need for any global mask to synchronize any thread, and we only need to loop over a number of groups to check 64 threads at a time per iteration. As such, tinfo's threads_want_rdv could be dropped. This was tested with 64 threads spread into 2 groups, running 64 tasks (from the debug dev command), 20 "show sess" (thread_isolate()), 20 "add server blah/blah" (thread_isolate()), and 20 "del server blah/blah" (thread_isolate_full()). The load remained very low (limited by external socat forks) and no stuck nor starved thread was found.	2022-07-01 19:15:15 +02:00
Willy Tarreau	ef422ced91	MEDIUM: thread: make stopping_threads per-group and add stopping_tgroups Stopping threads need a mask to figure who's still there without scanning everything in the poll loop. This means this will have to be per-group. And we also need to have a global stopping groups mask to know what groups were already signaled. This is used both to figure what thread is the first one to catch the event, and which one is the first one to detect the end of the last job. The logic isn't changed, though a loop is required in the slow path to make sure all threads are aware of the end. Note that for now the soft-stop still takes time for group IDs > 1 as the poller is not yet started on these threads and needs to expire its timeout as there's no way to wake it up. But all threads are eventually stopped.	2022-07-01 19:15:15 +02:00
Willy Tarreau	03f9b35114	MEDIUM: tinfo: add a dynamic thread-group context The thread group info is not sufficient to represent a thread group's current state as it's read-only. We also need something comparable to the thread context to represent the aggregate state of the threads in that group. This patch introduces ha_tgroup_ctx[] and tg_ctx for this. It's indexed on the group id and must be cache-line aligned. The thread masks that were global and that do not need to remain global were moved there (want_rdv, harmless, idle). Given that all the masks placed there now become group-specific, the associated thread mask (tid_bit) now switches to the thread's local bit (ltid_bit). Both are the same for nbtgroups 1 but will differ for other values. There's also a tg_ctx pointer in the thread so that it can be reached from other threads.	2022-07-01 19:15:15 +02:00
Willy Tarreau	22b2a24eb2	CLEANUP: thread: remove thread_sync_release() and thread_sync_mask This function was added in 2.0 when reworking the thread isolation mechanism to make it more reliable. However it if fundamentally incompatible with the full isolation mechanism provided by thread_isolate_full() since that one will wait for all threads to become idle while the former will wait for all threads to finish waiting, causing a deadlock. Given that it's not used, let's just drop it entirely before it gets used by accident.	2022-07-01 19:15:15 +02:00
Willy Tarreau	cce203aae5	MINOR: thread: add a new all_tgroups_mask variable to know about active tgroups In order to kill all_threads_mask we'll need to have an equivalent for the thread groups. The all_tgroups_mask does just this, it keeps one bit set per enabled group.	2022-07-01 19:15:15 +02:00
Willy Tarreau	c6cf64bb5e	MINOR: thread: use ltid_bit in ha_tkillall() Since commit `cc7a11ee3` ("MINOR: threads: set the tid, ltid and their bit in thread_cfg") we ought not use (1UL << thr) to get the group mask for thread <thr>, but (ha_thread_info[thr].ltid_bit). ha_tkillall() needs this.	2022-07-01 19:15:15 +02:00
Willy Tarreau	1e7f0d68b0	MINOR: clock: use ltid_bit in clock_report_idle() Since commit `cc7a11ee3` ("MINOR: threads: set the tid, ltid and their bit in thread_cfg") we ought not use (1UL << thr) to get the group mask for thread <thr>, but (ha_thread_info[thr].ltid_bit). clock_report_idle() needs this. This also implies not using all_threads_mask anymore but taking the mask from the tgroup since it becomes relative now.	2022-07-01 19:15:15 +02:00
Willy Tarreau	adc1f52c92	MINOR: wdt: use ltid_bit in wdt_handler() Since commit `cc7a11ee3` ("MINOR: threads: set the tid, ltid and their bit in thread_cfg") we ought not use (1UL << thr) to get the group mask for thread <thr>, but (ha_thread_info[thr].ltid_bit). wdt_handler() needs this.	2022-07-01 19:15:14 +02:00
Willy Tarreau	38d0712748	MINOR: debug: use ltid_bit in ha_thread_dump() Since commit `cc7a11ee3` ("MINOR: threads: set the tid, ltid and their bit in thread_cfg") we ought not use (1UL << thr) to get the group mask for thread <thr>, but (ha_thread_info[thr].ltid_bit). ha_thread_dump() needs this.	2022-07-01 19:15:14 +02:00
Willy Tarreau	377e37a80f	MINOR: tinfo: add the mask of enabled threads in each group In order to replace the global "all_threads_mask" we'll need to have an equivalent per group. Take this opportunity for calling it threads_enabled and make sure which ones are counted there (in case in the future we allow to stop some).	2022-07-01 19:15:14 +02:00
Willy Tarreau	60fe4a95a2	MINOR: tinfo: replace the tgid with tgid_bit in tgroup_info Now that the tgid is accessible from the thread, it's pointless to have it in the group, and it was only set but never used. However we'll soon frequently need the mask corresponding to the group ID and the risk of getting it wrong with the +1 or to shift 1 instead of 1UL is important, so let's store the tgid_bit there.	2022-07-01 19:15:14 +02:00
Willy Tarreau	66ad98a772	MINOR: tinfo: add the tgid to the thread_info struct At several places we're dereferencing the thread group just to catch the group number, and this will become even more required once we start to use per-group contexts. Let's just add the tgid in the thread_info struct to make this easier.	2022-07-01 19:15:14 +02:00
Willy Tarreau	e7475c8e79	MEDIUM: tasks/fd: replace sleeping_thread_mask with a TH_FL_SLEEPING flag Every single place where sleeping_thread_mask was still used was to test or set a single thread. We can now add a per-thread flag to indicate a thread is sleeping, and remove this shared mask. The wake_thread() function now always performs an atomic fetch-and-or instead of a first load then an atomic OR. That's cleaner and more reliable. This is not easy to test, as broadcast FD events are rare. The good way to test for this is to run a very low rate-limited frontend with a listener that listens to the fewest possible threads (2), and to send it only 1 connection at a time. The listener will periodically pause and the wakeup task will sometimes wake up on a random thread and will call wake_thread(): frontend test bind :8888 maxconn 10 thread 1-2 rate-limit sessions 5 Alternately, disabling/enabling a frontend in loops via the CLI also broadcasts such events, but they're more difficult to observe since this is causing connection failures.	2022-07-01 19:15:14 +02:00
Willy Tarreau	dce4ad755f	MEDIUM: thread: add a new per-thread flag TH_FL_NOTIFIED to remember wakeups Right now when an inter-thread wakeup happens, we preliminary check if the thread was asleep, and if so we wake the poller up and remove its bit from the sleeping mask. That's not very clean since the sleeping mask cannot be entirely trusted since a thread that's about to wake up will already have its sleeping bit removed. This patch adds a new per-thread flag (TH_FL_NOTIFIED) to remember that a thread was notified to wake up. It's cleared before checking the task lists last, so that new wakeups can be considered again (since wake_thread() is only used to notify about task wakeups and FD polling changes). This way we do not need to modify a remote thread's sleeping mask anymore. As such wake_thread() now only tests and sets the TH_FL_NOTIFIED flag but doesn't clear sleeping anymore.	2022-07-01 19:15:14 +02:00
Willy Tarreau	555c192d14	MINOR: poller: update_fd_polling: wake a random other thread When enabling an FD that's only bound to another thread, instead of always picking the first one, let's pick a random one. This is rarely used (enabling a frontend, or session rate-limiting period ending), and has greater chances of avoiding that some obscure corner cases could degenerate into a poorly distributed load.	2022-07-01 19:15:14 +02:00
Willy Tarreau	962e5ba72b	MEDIUM: polling: make update_fd_polling() not care about sleeping threads Till now, update_fd_polling() used to check if all the target threads were sleeping, and only then would wake an owning thread up. This causes several problems among which the need for the sleeping_thread_mask and the fact that by the time we wake one thread up, it has changed. This commit changes this by leaving it to wake_thread() to perform this check on the selected thread, since wake_thread() is already capable of doing this now. Concretely speaking, for updt_fd_polling() it will mean performing one computation of an ffsl() before knowing the sleeping status on a global FD state change (which is very rare and not important here, as it basically happens after relaxing a rate-limit (i.e. once a second at beast) or after enabling a frontend from the CLI); thus we don't care.	2022-07-01 19:15:14 +02:00
Willy Tarreau	058b2c1015	MINOR: poller: centralize poll return handling When returning from the polling syscall, all pollers have a certain dance to follow, made of wall clock updates, thread harmless updates, idle time management and sleeping mask updates. Let's have a centralized function to deal with all of this boring stuff: fd_leaving_poll(), and make all the pollers use it.	2022-07-01 19:15:14 +02:00
Willy Tarreau	bdcd32598f	MINOR: thread: only use atomic ops to touch the flags The thread flags are touched a little bit by other threads, e.g. the STUCK flag may be set by other ones, and they're watched a little bit. As such we need to use atomic ops only to manipulate them. Most places were already using them, but here we generalize the practice. Only ha_thread_dump() does not change because it's run under isolation.	2022-07-01 19:15:14 +02:00
Willy Tarreau	f3efef4d60	MINOR: thread: make wake_thread() take care of the sleeping threads mask Almost every call place of wake_thread() checks for sleeping threads and clears the sleeping mask itself, while the function is solely used for there. Let's move the check and the clearing of the bit inside the function itself. Note that updt_fd_polling() still performs the check because its rules are a bit different.	2022-07-01 19:15:14 +02:00
Willy Tarreau	3fdacdddaf	MEDIUM: queue: revert to regular inter-task wakeups Now that the inter-task wakeups are cheap, there's no point in using task_instant_wakeup() anymore when dequeueing tasks. The use of the regular task_wakeup() is sufficient there and will preserve a better fairness: the test that went from 40k to 570k RPS now gives 580k RPS (down from 585k RPS with previous commit). This essentially reverts commit `27fab1dcb` ("MEDIUM: queue: use tasklet_instant_wakeup() to wake tasks").	2022-07-01 19:15:14 +02:00
Willy Tarreau	319d136ff9	MEDIUM: task: use regular eb32 trees for the run queues Since we don't mix tasks from different threads in the run queues anymore, we don't need to use the eb32sc_ trees and we can switch to the regular eb32 ones. This uses cheaper lookup and insert code, and a 16-thread test on the queues shows a performance increase from 570k RPS to 585k RPS.	2022-07-01 19:15:14 +02:00
Willy Tarreau	c958c70ec8	MINOR: task: replace global_tasks_mask with a check for tree's emptiness This bit field used to be a per-thread cache of the result of the last lookup of the presence of a task for each thread in the shared cache. Since we now know that each thread has its own shared cache, a test of emptiness is now sufficient to decide whether or not the shared tree has a task for the current thread. Let's just remove this mask.	2022-07-01 19:15:14 +02:00
Willy Tarreau	da195e8aab	MINOR: task: remove grq_total and use rq_total instead grq_total was only used to know how many tasks were being queued in the global runqueue for stats purposes, and that was transferred to the per thread rq_total counter once assigned. We don't need this anymore since we know where they are, so let's just directly update rq_total and drop that one.	2022-07-01 19:15:14 +02:00
Willy Tarreau	b17dd6cc19	MEDIUM: task: replace the global rq_lock with a per-rq one There's no point having a global rq_lock now that we have one shared RQ per thread, let's have one lock per runqueue instead.	2022-07-01 19:15:14 +02:00
Willy Tarreau	6f78038d72	MEDIUM: task: move the shared runqueue to one per thread Since we only use the shared runqueue to put tasks only assigned to known threads, let's move that runqueue to each of these threads. The goal will be to arrange an N*(N-1) mesh instead of a central contention point. The global_rqueue_ticks had to be dropped (for good) since we'll now use the per-thread rqueue_ticks counter for both trees. A few points to note: - the rq_lock stlil remains the global one for now so there should not be any gain in doing this, but should this trigger any regression, it is important to detect whether it's related to the lock or to the tree. - there's no more reason for using the scope-based version of the ebtree now, we could switch back to the regular eb32_tree. - it's worth checking if we still need TASK_GLOBAL (probably only to delete a task in one's own shared queue maybe).	2022-07-01 19:15:14 +02:00
Willy Tarreau	a4fb79b4a2	MINOR: task: make rqueue_ticks atomic The runqueue ticks counter is per-thread and wasn't initially meant to be shared. We'll soon have to share it so let's make it atomic. It's only updated when waking up a task, and no performance difference was observed. It was moved in the thread_ctx struct so that it doesn't pollute the local cache line when it's later updated by other threads.	2022-07-01 19:15:14 +02:00
Willy Tarreau	fc5de15baa	CLEANUP: task: remove the now unused TASK_GLOBAL flag TASK_GLOBAL was exclusively used by task_unlink_rq(), as such it can be dropped.	2022-07-01 19:15:14 +02:00
Willy Tarreau	eed3911a54	MINOR: task: replace task_set_affinity() with task_set_thread() The latter passes a thread ID instead of a mask, making the code simpler.	2022-07-01 19:15:14 +02:00
Willy Tarreau	159e3acf5d	MEDIUM: task: remove TASK_SHARED_WQ and only use t->tid TASK_SHARED_WQ was set upon task creation and never changed afterwards. Thus if a task was created to run anywhere (e.g. a check or a Lua task), all its timers would always pass through the shared timers queue with a lock. Now we know that tid<0 indicates a shared task, so we can use that to decide whether or not to use the shared queue. The task might be migrated using task_set_affinity() but it's always dequeued first so the check will still be valid. Not only this removes a flag that's difficult to keep synchronized with the thread ID, but it should significantly lower the load on systems with many checks. A quick test with 5000 servers and fast checks that were saturating the CPU shows that the check rate increased by 20% (hence the CPU usage dropped by 17%). It's worth noting that run_task_lists() almost no longer appears in perf top now.	2022-07-01 19:15:14 +02:00
Willy Tarreau	3b7a19c2a6	MINOR: applet: always use task_new_on() on applet creation Now that task_new_on() supports negative numbers, there's no need for checking the thread value nor falling back to task_new_anywhere().	2022-07-01 19:15:14 +02:00
Willy Tarreau	cb8542755e	MEDIUM: applet: only keep appctx_new_*() and drop appctx_new() This removes the mask-based variant so that from now on the low-level function becomes appctx_new_on() and it takes either a thread number or a negative value for "any thread". This way we can use task_new_on() and task_new_anywhere() instead of task_new() which will soon disappear.	2022-07-01 19:15:14 +02:00
Willy Tarreau	0ad00befc1	CLEANUP: task: remove thread_mask from the struct task It was not used anymore since everything moved to ->tid, so let's remove it.	2022-07-01 19:15:14 +02:00
Willy Tarreau	c44d08ebc4	MAJOR: task: replace t->thread_mask with 1<<t->tid when thread mask is needed At a few places where the task's thread mask. Now we know that it's always either one bit or all bits of all_threads_mask, so we can replace it with either 1<<tid or all_threads_mask depending on what's expected. It's worth noting that the global_tasks_mask is still set this way and that it's reaching its limits. Similarly, the task_new() API would deserve an update to stop using a thread mask and use a thread number instead. Similarly, task_set_affinity() should be updated to directly take a thread number. At this point the task's thread mask is not used anymore.	2022-07-01 19:15:14 +02:00
Willy Tarreau	29ffe26733	MAJOR: task: use t->tid instead of ffsl(t->thread_mask) to take the thread ID At several places we need to figure the ID of the first thread allowed to run a task. Till now this was performed using my_ffsl(t->thread_mask) but since we now have the thread ID stored into the task, let's use it instead. This is tagged major because it starts to assume that tid<0 is strictly equivalent to atleast2(thread_mask), and that as such, among the allowed threads are the current one.	2022-07-01 19:15:14 +02:00
Willy Tarreau	6ef52f4479	MEDIUM: task: add and preset a thread ID in the task struct The tasks currently rely on a mask but do not have an assigned thread ID, contrary to tasklets. However, in practice they're either running on a single thread or on any thread, so that it will be worth simplifying all this in order to ease the transition to the thread groups. This patch introduces a "tid" field in the task struct, that's either the number of the thread the task is attached to, or a negative value if the task is not bound to a thread, (i.e. its mask is all_threads_mask). The new ID is only set and updated but not used yet.	2022-07-01 19:15:14 +02:00
Willy Tarreau	8e5c53a6c9	MINOR: debug: remove mask support from "debug dev sched" The thread mask will not be used anymore, instead the thread id only is used. Interestingly it was already implemented in the parsing but not used. The single/multi thread argument is not needed anymore since it's sufficient to pass tid<0 to get a multi-threaded task/tasklet. This is in preparation for the removal of the thread_mask in tasks as only this debug code was using it!	2022-07-01 19:15:14 +02:00
Emeric Brun	7d392a592d	BUG/MEDIUM: ssl/fd: unexpected fd close using async engine Before 2.3, after an async crypto processing or on session close, the engine async file's descriptors were removed from the fdtab but not closed because it is the engine which has created the file descriptor, and it is responsible for closing it. In 2.3 the fd_remove() call was replaced by fd_stop_both() which stops the polling but does not remove the fd from the fdtab and the fd remains indefinitively in the fdtab. A simple replacement by fd_delete() is not a valid fix because fd_delete() removes the fd from the fdtab but also closes the fd. And the fd will be closed twice: by the haproxy's core and by the engine itself. Instead, let's set FD_DISOWN on the FD before calling fd_delete() which will take care of not closing it. This patch must be backported on branches >= 2.3, and it relies on this previous patch: MINOR: fd: add a new FD_DISOWN flag to prevent from closing a deleted FD As mentioned in the patch above, a different flag will be needed in 2.3.	2022-07-01 17:41:40 +02:00
Emeric Brun	f41a3f6762	MINOR: fd: add a new FD_DISOWN flag to prevent from closing a deleted FD Some FDs might be offered to some external code (external libraries) which will deal with them until they close them. As such we must not close them upon fd_delete() but we need to delete them anyway so that they do not appear anymore in the fdtab. This used to be handled by fd_remove() before 2.3 but we don't have this anymore. This patch introduces a new flag FD_DISOWN to let fd_delete() know that the core doesn't own the fd and it must not be closed upon removal from the fd_tab. This way it's totally unregistered from the poller but still open. This patch must be backported on branches >= 2.3 because it will be needed to fix a bug affecting SSL async. it should be adapted on 2.3 because state flags were stored in a different way (via bits in the structure).	2022-07-01 17:41:40 +02:00
Amaury Denoyelle	6befccd8a1	BUG/MINOR: mux-quic: do not signal FIN if gap in buffer Adjust FIN signal on Rx path for the application layer : ensure that the receive buffer has no gap. Without this extra condition, FIN was signalled as soon as the STREAM frame with FIN was received, even if we were still waiting to receive missing offsets. This bug could have lead to incomplete requests read from the application protocol. However, in practice this bug has very little chance to happen as the application layer ensures that the demuxed frame length is equivalent to the buffer data size. The only way to happen is if to receive the FIN STREAM as the H3 demuxer is still processing on a frame which is not the last one of the stream. This must be backported up to 2.6. The previous patch on ncbuf is required for the newly defined function ncb_is_fragmented(). MINOR: ncbuf: implement ncb_is_fragmented()	2022-07-01 15:55:32 +02:00
Amaury Denoyelle	e0a92a7e56	MINOR: ncbuf: implement ncb_is_fragmented() Implement a new status function for ncbuf. It allows to quickly report if a buffer contains data in a fragmented way, i.e. with gaps in between or at start of the buffer. To summarize, a buffer is considered as non-fragmented in the following cases : - a null or empty buffer - a full buffer - a buffer containing exactly one data block at the beginning, following by a gap until the end.	2022-07-01 15:54:23 +02:00
Amaury Denoyelle	36d4b5e31d	CLEANUP: mux-quic: adjust comment on qcs_consume() Since a previous refactoring, application protocol layer is not require anymore to call qcs_consume(). This function is now automatically used by the MUX itself.	2022-07-01 14:46:24 +02:00
Frédéric Lécaille	67fda16742	CLEANUP: h2: Typo fix in h2_unsubcribe() traces Very minor modification for the traces of this function.	2022-06-30 14:34:32 +02:00
Frédéric Lécaille	1b0707f3e7	MINOR: quic: Improvements for the datagrams receipt First we add a loop around recfrom() into the most low level I/O handler quic_sock_fd_iocb() to collect as most as possible datagrams before during its tasklet wakeup with a limit: we recvfrom() at most "maxpollevents" datagrams. Furthermore we add a local task list into the datagram handler quic_lstnr_dghdlr() which is passed to the first datagrams parser qc_lstnr_pkt_rcv(). This latter parser only identifies the connection associated to the datagrams then wakeup the highest level packet parser I/O handlers (quic_conn.*io_cb()) after it is done, thanks to the call to tasklet_wakeup_after() which replaces from now on the call to tasklet_wakeup(). This should reduce drastically the latency and the chances to fulfil the RX buffers at the QUIC connections level as reported in GH #1737 by Tritan. These modifications depend on this commit: "MINOR: task: Add tasklet_wakeup_after()" Must be backported to 2.6 with the previous commit.	2022-06-30 14:34:27 +02:00
Frédéric Lécaille	45a16295e3	MINOR: quic: Add new stats counter to diagnose RX buffer overrun Remove the call to qc_list_all_rx_pkts() which print messages on stderr during RX buffer overruns and add a new counter for the number of dropped packets because of such events. Must be backported to 2.6	2022-06-30 14:24:04 +02:00
Frédéric Lécaille	95a8dfb4c7	BUG/MINOR: quic: Dropped packets not counted (with RX buffers full) When the connection RX buffer is full, the received packets are dropped. Some of them were not taken into an account by the ->dropped_pkt counter. This very simple patch has no impact at all on the packet handling workflow. Must be backported to 2.6.	2022-06-30 14:24:04 +02:00
Frédéric Lécaille	ad548b54a7	MINOR: task: Add tasklet_wakeup_after() We want to be able to schedule a tasklet onto a thread after the current tasklet is done. What we have to do is to insert this tasklet at the head of the thread task list. Furthermore, we would like to serialize the tasklets. They must be run in the same order as the order in which they have been scheduled. This is implemented passing a list of tasklet as parameter (see <head> parameters) which must be reused for subsequent calls. _tasklet_wakeup_after_on() is implemented to accomplish this job. tasklet_wakeup_after_on() and tasklet_wake_after() are only wrapper macros around _tasklet_wakeup_after_on(). tasklet_wakeup_after_on() does exactly the same thing as _tasklet_wakeup_after_on() without having to pass the filename and line in the filename as parameters (usefull when DEBUG_TASK is enabled). tasklet_wakeup_after() hides also the usage of the thread parameter which is <tl> tasklet thread ID.	2022-06-30 14:24:04 +02:00
Amaury Denoyelle	a7a4c80ade	MINOR: qpack: properly handle invalid dynamic table references Return QPACK_DECOMPRESSION_FAILED error code when dealing with dynamic table references. This is justified as for now haproxy does not implement dynamic table support and advertizes a zero-sized table. The H3 calling function will thus reuse this code in a CONNECTION_CLOSE frame, in conformance with the QPACK RFC. This proper error management allows to remove obsolete ABORT_NOW guards.	2022-06-30 11:51:06 +02:00
Amaury Denoyelle	46e992d795	BUG/MINOR: qpack: abort on dynamic index field line decoding This is a complement to partial fix from commit `debaa04f9e` BUG/MINOR: qpack: abort on dynamic index field line decoding The main objective is to fix coverity report about usage of uninitialized variable when receiving dynamic table references. These references are invalid as for the moment haproxy advertizes a 0-sized dynamic table. An ABORT_NOW clause is present to catch this. A following patch will clean up this in order to properly handle QPACK errors with CONNECTION_CLOSE. This should fix github issue #1753. No need to backport as this was introduced in the current dev branch.	2022-06-30 11:51:06 +02:00
Amaury Denoyelle	2bc47863ec	MINOR: h3: handle errors on HEADERS parsing/QPACK decoding Emit a CONNECTION_CLOSE if HEADERS parsing function returns an error. This is useful to remove previous ABORT_NOW guards. For the moment, the whole connection is closed. In the future, it may be justified to only reset the faulting stream in some cases. This requires the implementation of RESET_STREAM emission.	2022-06-30 11:51:06 +02:00
Amaury Denoyelle	055de23b7d	BUG/MINOR: qpack: fix build with QPACK_DEBUG The local variable 't' was renamed 'static_tbl'. Fix its name in the qpack_debug_printf() statement which is activated only with QPACK_DEBUG mode. No need to backport as this was introduced in current dev branch.	2022-06-30 10:15:58 +02:00
Christopher Faulet	2b6777021d	MEDIUM: bwlim: Add support of bandwith limitation at the stream level This patch adds a filter to limit bandwith at the stream level. Several filters can be defined. A filter may limit incoming data (upload) or outgoing data (download). The limit can be defined per-stream or shared via a stick-table. For a given stream, the bandwith limitation filters can be enabled using the "set-bandwidth-limit" action. A bandwith limitation filter can be used indifferently for HTTP or TCP stream. For HTTP stream, only the payload transfer is limited. The filter is pretty simple for now. But it was designed to be extensible. The current design tries, as far as possible, to never exceed the limit. There is no burst.	2022-06-24 14:06:26 +02:00
Frédéric Lécaille	628e89cfae	BUILD: quic+h3: 32-bit compilation errors fixes In GH #1760 (which is marked as being a feature), there were compilation errors on MacOS which could be reproduced in Linux when building 32-bit code (-m32 gcc option). Most of them were due to variables types mixing in QUIC_MIN macro or using size_t type in place of uint64_t type. Must be backported to 2.6.	2022-06-24 12:13:53 +02:00
Fr�d�ric L�caille	2bed1f166e	BUG/MAJOR: quic: Big RX dgrams leak with POST requests This previous commit: "BUG/MAJOR: Big RX dgrams leak when fulfilling a buffer" partially fixed an RX dgram memleak. There is a missing break in the loop which looks for the first datagram attached to an RX buffer dgrams list which may be reused (because consumed by the connection thread). So when several dgrams were consumed by the connection thread and are present in the RX buffer list, some are leaked because never reused for ever. They are removed for their list. Furthermore, as commented in this patch, there is always at least one dgram object attached to an RX dgrams list, excepted the first time we enter this I/O handler function for this RX buffer. So, there is no need to use a loop to lookup and reuse the first datagram in an RX buffer dgrams list. This isssue was reproduced with quiche client with plenty of POST requests (100000 streams): cargo run --bin quiche-client -- https://127.0.0.1:8080/helloworld.html --no-verify -n 100000 --method POST --body /var/www/html/helloworld.html and could be reproduce with GET request. This bug was reported by Tristan in GH #1749. Must be backported to 2.6.	2022-06-23 21:57:09 +02:00
Fr�d�ric L�caille	19ef6369b5	BUG/MAJOR: quic: Big RX dgrams leak when fulfilling a buffer When entering quic_sock_fd_iocb() I/O handler which is responsible of recvfrom() datagrams, the first thing which is done it to try to reuse a dgram object containing metadata about the received datagrams which has been consumed by the connection thread. If this object could not be used for any reason, so when we "goto out" of this function, we must release the memory allocated for this objet, if not it will leak. Most of the time, this happened when we fulfilled a buffer as reported in GH #1749 by Tristan. This is why we added a pool_free() call just before the out label. We mark <new_dgram> as NULL when it successfully could be used. Thank you for Tristan and Willy for their participation on this issue. Must be backported to 2.6.	2022-06-23 20:40:01 +02:00
Fr�d�ric L�caille	0c535683ee	BUG/MINOR: quic: Wrong reuse of fulfilled dgram RX buffer After having fulfilled a buffer, then marked it as full, we must consume the remaining space. But to do that, and not to erase the already existing data, we must check there is not remaining data in after the tail of the buffer (between the tail and the head). This is done adding a condition to test that adding the number of bytes from the remaining contiguous space to the tail does not pass the wrapping postion in the buffer. Must be backported to 2.6.	2022-06-23 20:39:19 +02:00
Willy Tarreau	27061cd144	MEDIUM: debug: improve DEBUG_MEM_STATS to also report pool alloc/free Sometimes using "debug dev memstats" can be frustrating because all pool allocations are reported through pool-os.h and that's all. But in practice there's nothing wrong with also intercepting pool_alloc, pool_free and pool_zalloc and report their call counts and locations, so that's what this patch does. It only uses an alternate set of macroes for these 3 calls when DEBUG_MEM_STATS is defined. The outputs are reported as P_ALLOC (for both pool_malloc() and pool_zalloc()) and P_FREE (for pool_free()).	2022-06-23 11:58:01 +02:00
Willy Tarreau	b8dec4a01a	CLEANUP: pool/tree-wide: remove suffix "_pool" from certain pool names A curious practise seems to have started long ago and contaminated various code areas, consisting in appending "_pool" at the end of the name of a given pool. That makes no sense as the name is only used to name the pool in diags such as "show pools", and since names are truncated there, this adds some confusion when analysing the dump outputs. Let's just clean all of them at once. there were essentially in SSL and QUIC.	2022-06-23 11:49:09 +02:00
Willy Tarreau	47af317389	BUG/MINOR: stream: only free the req/res captures when set There's a subtle bug in stream_free() when releasing captures. The pools may be NULL when no capture is defined, and the calls to pool_free() are inconditional. The only reason why this doesn't cause trouble is because the pointer to be freed is always NULL in this case and we don't go further down the chain. That's particularly ugly and it complicates debugging, so let's only call these ones when the pointers are set. There's no impact on running code, it only fools those trying to debug pools manually. There's no need to backport it though it unless it helps for debugging sessions.	2022-06-23 11:49:09 +02:00
Christopher Faulet	aa55640b8c	MINOR: freq_ctr: Add a function to get events excess over the current period freq_ctr_overshoot_period() function may be used to retrieve the excess of events over the current period for a givent frequency counter, ignoring the history. It is a way compare the "current rate" (the number of events over the current period) to a given rate and estimate the excess of events. It may be used to safely add new events, especially at the begining of the current period for a frequency counter with large period. This way, it is possible to smoothly add events during the whole period without quickly consuming all the quota at the beginning of the period and waiting for the next one to be able to add new events.	2022-06-22 18:33:27 +02:00
Christopher Faulet	dbbdb25f1c	BUG/MINOR: http-fetch: Use integer value when possible in "method" sample fetch Because of the previous fix, if the HTTP parsing is performed when the "method" sample fetch is called, we always rely on the string representation of the request method. Indeed, if no parsing was performed when the "method" sample fetch is called, the transaction method is HTTP_METH_OTHER because it was just initialized. However, without this patch, in this case, we always retrieve the method by reading the request start-line. Now, when the method is HTTP_METH_OTHER, we systematically try to parse the request but the method is tested once again after the parsing to be able to use the integer representation when possible. This patch must be backported as far as 2.0.	2022-06-22 17:50:54 +02:00
Christopher Faulet	5eb67f5d74	BUG/MINOR: http-ana: Set method to HTTP_METH_OTHER when an HTTP txn is created This patch is required to fix "method" sample fetch. But it make sense to initialize the method of an HTTP transaction to HTTP_METH_OTHER. This way, before the request parsing, the method is considered as unknown except if we are able to retrieve the request start-line. It is especially important for TCP streams. About the "method" sample fetch, this patch is a way to be sure no random method is returned when the sample fetch is used on a TCP stream before any HTTP parsing. This patch must be backported as far as 2.0.	2022-06-22 17:50:54 +02:00
Frédéric Lécaille	77ac6f5667	BUG/MINOR: quic: Missing acknowledgments for trailing packets This bug was revealed by key_update QUIC tracker test. During this test, after the handshake has succeeded, the client sends a 1-RTT packet containing only a PING frame. On our side, we do not acknowledge it, because we have no ack-eliciting packet to send. This is not correct. We must acknowledge all the ack-eliciting packets unconditionally. But we must not send too much ACK frames: every two packets since we have sent an ACK frame. This is the test (nb_aepkts_since_last_ack >= QUIC_MAX_RX_AEPKTS_SINCE_LAST_ACK) which ensure this is the case. But this condition must be checked at the very last time, when we are building a packet and when an acknowledgment is required. This is not to qc_may_build_pkt() to do that. This boolean function decides if we have packets to send. It must return "true" if there is no more ack-eliciting packets to send and if an acknowledgement is required. We must also add a "force_ack" parameter to qc_build_pkt() to force the acknowledments during the handshakes (for each packet). If not, they will be sent every two packets. This parameter must be passed to qc_do_build_pkt() and be checked alongside the others conditions when building a packet to decide to add an ACK frame or not to this packet. Must be backported to 2.6.	2022-06-22 15:47:52 +02:00
Remi Tricot-Le Breton	1bad7db4a1	BUG/MINOR: ssl: Do not look for key in extra files if already in pem A bug was introduced by commit `9bf3a1f67e` "BUG/MINOR: ssl: Fix crash when no private key is found in pem". If a private key is already contained in a pem file, we will still look for a .key file and load its private key if it exists when we should not. This patch should be backported to all branches where the original fix was backported (all the way to 2.2).	2022-06-22 10:45:47 +02:00
Willy Tarreau	d543ae0e68	BUILD: ssl_ckch: fix "maybe-uninitialized" build error on gcc-9.4 + ARM As reported in issue #1755, gcc-9.3 and 9.4 emit a "maybe-uninitialized" warning in cli_io_handler_commit_cafile_crlfile() because it sees that when the "path" variable is not set, we're jumping to the error label inside the loop but cannot see that the new state will avoid the places where the value is used. Thus it's a false positive but a difficult one. Let's just preset the value to NULL to make it happy. This was introduced in 2.7-dev by commit `ddc8e1cf8` ("MINOR: ssl_ckch: Simplify I/O handler to commit changes on CA/CRL entry"), thus no backport is needed for now.	2022-06-22 05:45:34 +02:00
Willy Tarreau	c7a8a3c7bd	MINOR: intops: add a function to return a valid bit position from a mask Sometimes we need to be able to signal one thread among a mask, without caring much about which bit will be picked. At the moment we use ffsl() for this but this sometimes results in imbalance at certain specific places where the same first thread in a set is always the same one that is selected. Another approach would consist in using the rank finding function but it requires a popcount and a setup phase, and possibly a modulo operation depending on the popcount, which starts to be very expensive. Here we take a different approach. The idea is an input bit value is passed, from 0 to LONGBITS-1, and that as much as possible we try to pick the bit matching it if it is set. Otherwise we look at a mirror position based on a decreasing power of two, and jump to the side that still has bits left. In 6 iterations it ends up spotting one bit among 64 and the operations are very cheap and optimizable. This method has the benefit that we don't care where the holes are located in the mask, thus it shows a good distribution of output bits based on the input ones. A long-time test shows an average of 16 cycles, or ~4ns per lookup at 3.8 GHz, which is about twice as fast as using the rank finding function. Just like for that one, the code was stored into tools.c since we don't have a C file for intops.	2022-06-21 20:29:57 +02:00
William Lallemand	0a012aa16b	BUG/MEDIUM: mworker: use default maxconn in wait mode In bug #1751, it was reported that haproxy is consumming too much memory since the 2.4 version. This is because of a change in the master, which loses completely its configuration in wait mode, and lose its maxconn. Without the maxconn, haproxy will try to compute one itself, and will allocate RAM consequently, too much in our case. Which means the master will have a too high maxconn and too much RAM allocated. The patch fixes the issue by setting the maxconn to the default value when re-executing the master in wait mode. Must be backported as far as 2.5.	2022-06-21 14:22:49 +02:00
Frédéric Lécaille	4f5777a415	MINOR: quic: Dump version_information transport parameter Implement quic_tp_version_info_dump() to dump such a transport parameter (only remote). Call it from quic_transport_params_dump() which dump all the transport parameters. Can be backported to 2.6 as it's useful for debugging.	2022-06-21 11:07:39 +02:00
Frédéric Lécaille	57bddbcbbb	BUG/MINOR: quic: Acknowledgement must be forced during handshake All packets received during hanshakes must be acknowledged asap. This was not the case for Handshake packets received. At this time, this had no impact because the client has often only one Handshake packet to send and last handshake to be sent on our side always embeds an HANDSHAKE_DONE frame which leads the client to consider it has no more handshake packet to send. Add <force_ack> to qc_may_build_pkt() to force an ACK frame to be sent. Set this parameter to 1 when sending packets from Initial or Handshake packet number spaces, 0 when sending only Application level packet. Must be backported to 2.6.	2022-06-21 11:00:16 +02:00
William Lallemand	cb6c5f4683	BUG/MEDIUM: ssl/cli: crash when crt inserted into a crt-list The crash occures when the same certificate which is used on both a server line and a bind line is inserted in a crt-list over the CLI. This is quite uncommon as using the same file for a client and a server certificate does not make sense in a lot of environments. This patch fixes the issue by skipping the insertion of the SNI when no bind_conf is available in the ckch_inst. Change the reg-test to reproduce this corner case. Should fix issue #1748. Must be backported as far as 2.2. (it was previously in ssl_sock.c)	2022-06-20 17:27:49 +02:00
Amaury Denoyelle	debaa04f9e	BUG/MINOR: qpack: abort on dynamic index field line decoding Add an ABORT_NOW() clause if indexed field line referred to the dynamic table. This is required as current haproxy QPACK implementation does not support the dynamic table. Note that this should not happen as haproxy explicitely advertizes a null-sized dynamic table to the other peer. This shoud fix github issue #1753. No need to backport as this was introduced by commit `b666c6b26e` MINOR: qpack: improve decoding function	2022-06-20 15:56:01 +02:00
Amaury Denoyelle	23f908ccd6	BUG/MINOR: quic: free rejected Rx packets Free Rx packet in the datagram handler if the packet was not taken in charge by a quic_conn instance. This is reflected by the packet refcount which is null. A packet can be rejected for a variety of reasons. For example, failed decryption, no Initial token and Retry emission or for datagram null padding. This patch should resolve the Rx packets memory leak observed via "show pools" with the previous commit `2c31e12936` BUG/MINOR: quic: purge conn Rx packet list on release This specific memory leak instance was reproduced using quiche client which uses null datagram padding. This should partially resolve github issue #1751. It must be backported up to 2.6.	2022-06-20 15:04:07 +02:00
Amaury Denoyelle	2c31e12936	BUG/MINOR: quic: purge conn Rx packet list on release When releasing a quic_conn instance, free all remaining Rx packets in quic_conn.rx.pkt_list. This partially fixes a memory leak on Rx packets which can be observed after several QUIC connections establishment. This should partially resolve github issue #1751. It must be backported up to 2.6.	2022-06-20 14:59:52 +02:00
Frédéric Lécaille	483499dc22	BUG/MINOR: quic_stats: Duplicate "quic_streams_data_blocked_bidi" field name As reported by broxio in GH #1757, there was a duplication field name for "quic_streams_data_blocked_bidi", due to a copy and paste without renaming I guess. Must be backported to 2.6.	2022-06-20 14:57:19 +02:00
Frédéric Lécaille	2aebaa49b1	BUG/MINOR: quic: Unexpected half open connection counter wrapping This counter must be incremented only one time by connection and decremented as soon as the handshake has failed or succeeded. This is a gauge. Under certain conditions this counter could be decremented twice. For instance after having received a TLS alert, then upon SSL_do_handshake() failure. To stop having to deal to all the current combinations which can lead to such a situation (and the next to come), add a connection flag to denote if this counter has been already decremented for a connection. So, this counter must be decremented only if this flag has not been already set. Must be backported up to 2.6.	2022-06-20 14:57:09 +02:00
Frédéric Lécaille	b1cb958581	BUILD: quic: Wrong HKDF label constant variable initializations Non constant expressions were used to initialize constant variables leading to such compilation errors: src/xprt_quic.c:66:3: error: initializer element is not a constant expression .key_label_len = strlen(QUIC_HKDF_KEY_LABEL_V1), Reproduced with CC=gcc-4.9 compilation option. Fix using macros for each HKDF label.	2022-06-20 14:50:19 +02:00
Willy Tarreau	177aed56dc	MEDIUM: debug: detect redefinition of symbols upon dlopen() In order to better detect the danger caused by extra shared libraries which replace some symbols, upon dlopen() we now compare a few critical symbols such as malloc(), free(), and some OpenSSL symbols, to see if the loaded library comes with its own version. If this happens, a warning is emitted and TAINTED_REDEFINITION is set. This is important because some external libs might be linked against different libraries than the ones haproxy was linked with, and most often this will end very badly (e.g. an OpenSSL object is allocated by haproxy and freed by such libs). Since the main source of dlopen() calls is the Lua lib, via a "require" statement, it's worth trying to show a Lua call trace when detecting a symbol redefinition during dlopen(). As such we emit a Lua backtrace if Lua is detected as being in use.	2022-06-19 17:58:32 +02:00
Willy Tarreau	40dde2d5c1	MEDIUM: debug: add a tainted flag when a shared library is loaded Several bug reports were caused by shared libraries being loaded by other libraries or some Lua code. Such libraries could define alternate symbols or include dependencies to alternate versions of a library used by haproxy, making it very hard to understand backtraces and analyze the issue. Let's intercept dlopen() and set a new TAINTED_SHARED_LIBS flag when it succeeds, so that "show info" makes it visible that some external libs were added. The redefinition is based on the definition of RTLD_DEFAULT or RTLD_NEXT that were previously used to detect that dlsym() is usable, since we need it as well. This should be sufficient to catch most situations.	2022-06-19 17:58:32 +02:00
Willy Tarreau	0b7b639d7e	MINOR: hlua: add a new hlua_show_current_location() function This function may be used to try to show where some Lua code is currently being executed. It tries hard to detect the initialization phase, both for the global and the per-thread states, and for runtime states. This intends to be used by error handlers to provide the users with indications about what Lua code was being executed when the error triggered.	2022-06-19 17:58:32 +02:00
Willy Tarreau	5c143404ea	MINOR: hlua: don't dump empty entries in hlua_traceback() Calling hlua_traceback() sometimes reports empty entries looking like: [C]: ? These ones correspond to certain internal C functions of the Lua library, but they do not provide any information and even complicate the interpretation of the dump. Better just skip them.	2022-06-19 17:58:32 +02:00
Christopher Faulet	a892b7f15f	BUG/MINOR: log: Properly test connection retries to fix dontlog-normal option The commit `731c8e6cf` ("MINOR: stream: Simplify retries counter calculation") introduced a regression. It broke the dontlog-normal option because the test on the connection retries counter was not updated accordingly. This patch should fix the issue #1754. It must be backported to 2.6.	2022-06-17 14:53:21 +02:00
Christopher Faulet	82af3c684d	CLEANUP: stconn: Don't expect to have no sedesc on detach The stream connector must always have a defined sedesc. So there is no reason to test it when the stconn is detached from the endpoint.	2022-06-17 13:25:02 +02:00
Christopher Faulet	9b8d7a11c0	MINOR: stream: Rely on stconn flags to abort stream destructive upgrade On destructive connection upgrade, instead of using the new mux name to abort the old stream, we can relay on the stream connector flags. If it is detached after the upgrade, it means the stream will not be resused by the new mux and it must be aborted. This patch may be backported to 2.6.	2022-06-17 13:25:02 +02:00
Christopher Faulet	b68f77d626	BUG/MEDIUM: stream: Properly handle destructive client connection upgrades When the protocol is changed for a client connection at the stream level (from TCP to H1/H2), there are two cases. The stream may be reused or not. The first case, when the stream is reused is working. The second one is buggy since the conn-stream refactoring and leads to a crash. In this case, the new mux don't reuse the stream. It must be silently aborted. However, its front stream connector is still referencing the connection. So it must be detached. But it must be performed in two stages, to be sure to not loose the context for the upgrade and to be able to rollback on error. So now, before the upgrade, we prepare to detach the stconn and it is finally detached if the upgrade succeeds. There is a trick here. Because we pretend the stconn is detached but its state is preserved. This patch must be backported to 2.6.	2022-06-17 13:25:02 +02:00
Willy Tarreau	9b3aa63df7	BUG/MINOR: task: fix thread assignment in tasklet_kill() tasklet_kill() was introduced in 2.5-dev4 with commit `7b368339a` ("MEDIUM: task: implement tasklet kill"), but a comparison error there makes tasklets killed on thread 1 assigned to the killing thread. Fortunately, the function was finally not used so there's no harm right now, hence the minor tag, but this must be fixed and backported in case a later fix relies on it. This should be backported to 2.5.	2022-06-16 18:17:44 +02:00
Frédéric Lécaille	e06f7459fa	CLEANUP: quic: Remove any reference to boringssl I do not think we will support boringssl for QUIC soon ;)	2022-06-16 15:58:48 +02:00
Frédéric Lécaille	301425b880	MEDIUM: quic: Compatible version negotiation implementation (draft-08) At this time haproxy supported only incompatible version negotiation feature which consists in sending a Version Negotiation packet after having received a long packet without compatible value in its version field. This version value is the version use to build the current packet. This patch does not modify this behavior. This patch adds the support for compatible version negotiation feature which allows endpoints to negotiate during the first flight or packets sent by the client the QUIC version to use for the connection (or after the first flight). This is done thanks to "version_information" parameter sent by both endpoints. To be short, the client offers a list of supported versions by preference order. The server (or haproxy listener) chooses the first version it also supported as negotiated version. This implementation has an impact on the tranport parameters handling (in both direcetions). Indeed, the server must sent its version information, but only after received and parsed the client transport parameters). So we cannot encode these parameters at the same time we instantiated a new connection. Add QUIC_TP_DRAFT_VERSION_INFORMATION(0xff73db) new transport parameter. Add tp_version_information new C struct to handle this new parameter. Implement quic_transport_param_enc_version_info() (resp. quic_transport_param_dec_version_info()) to encode (resp. decode) this parameter. Add qc_conn_finalize() which encodes the transport parameters and configure the TLS stack to send them. Add ->negotiated_ictx quic_conn C struct new member to store the Initial QUIC TLS context for the negotiated version. The Initial secrets derivation is version dependent. Rename ->version to ->original_version and add ->negotiated_version to this C struct to reflect the QUIC-VN RFC denomination. Modify most of the QUIC TLS API functions to pass a version as parameter. Export the QUIC version definitions to be reused at least from quic_tp.c (transport parameters. Move the token check after the QUIC connection lookup. As this is the original version which is sent into a Retry packet, and because this original version is stored into the connection, we must check the token after having retreived this connection. Add packet version to traces. See https://datatracker.ietf.org/doc/html/draft-ietf-quic-version-negotiation-08 for more information about this new feature.	2022-06-16 15:58:48 +02:00
Frédéric Lécaille	e17bf77218	MINOR: quic: Released QUIC TLS extension for QUIC v2 draft This is not clear at all how to distinguish a QUIC draft version number from a released one. And among these QUIC draft versions, which one must use the draft QUIC TLS extension. According to the QUIC implementations which support v2 draft, the TLS extension (transport parameters) to be used is the released one (TLS_EXTENSION_QUIC_TRANSPORT_PARAMETERS). As the unique QUIC draft version we support is 0xff00001d and as at this time the unique version with 0xff as most significant byte is this latter which must use the draft TLS extension, we select the draft TLS extension (TLS_EXTENSION_QUIC_TRANSPORT_PARAMETERS_DRAFT) only for such versions with 0xff as most signification byte.	2022-06-16 14:56:24 +02:00
Frédéric Lécaille	86845c5171	MEDIUM: quic: Add QUIC v2 draft support This is becoming difficult to handle the QUIC TLS related definitions which arrive with a QUIC version (draft or not). So, here we add quic_version C struct which does not define only the QUIC version number, but also the QUIC TLS definitions which depend on a QUIC version. Modify consequently all the QUIC TLS API to reuse these definitions through new quic_version C struct. Implement quic_pkt_type() function which return a packet type (0 up to 3) depending on the QUIC version number. Stop harding the Retry packet first byte in send_retry(): this is not more possible because the packet type field depends on the QUIC version. Also modify quic_build_packet_long_header() for the same reason: the packet type depends on the QUIC version. Add a quic_version C struct member to quic_conn C struct. Modify qc_lstnr_pkt_rcv() to set this member asap. Remove the version member from quic_rx_packet C struct: a packet is attached asap to a connection (or dropped) which is the unique object which should store the QUIC version. Modify qc_pkt_is_supported_version() to return a supported quic_version C struct from a version number. Add Initial salt for QUIC v2 draft (initial_salt_v2_draft).	2022-06-16 14:56:24 +02:00
Frédéric Lécaille	ea0ec27eb4	MINOR: quic: Parse long packet version from qc_parse_hd_form() This is to prepare the support for QUIC v2 version. The packet type depends on the version. So, we must parse early enough the version before defining the type of each packet.	2022-06-16 14:56:24 +02:00
Frédéric Lécaille	3f96a0a4c1	MINOR: quic: Add several nonce and key definitions for Retry tag The nonce and keys used to cipher the Retry tag depend on the QUIC version. Add these definitions for 0xff00001d (draft-29) and v2 QUIC version. At least draft-29 is useful for QUIC tracker tests with "quic-force-retry" enabled on haproxy side. Validated with -v 0xff00001d ngtcp2 option. Could not validate the v2 nonce and key at this time because not supported.	2022-06-16 14:56:24 +02:00
Frédéric Lécaille	01d515e013	BUG/MINOR: quic: Stop hardcoding Retry packet Version field Use the same version as the one received. This is safe because the version is treated before anything else sending a Version packet. Must be backported to 2.6.	2022-06-16 14:56:24 +02:00
Amaury Denoyelle	fa7fadca19	BUG/BUILD: h3: fix wrong label name A pretty ugly mistake introduced recently with an invalid goto statement which prevents QUIC compilation on haproxy. This must be backported on 2.6 as a complement to `60ef19f137` BUG/MINOR: h3/qpack: deal with too many headers	2022-06-15 15:52:27 +02:00
Amaury Denoyelle	b666c6b26e	MINOR: qpack: improve decoding function Adjust decoding loop by using temporary ist for header name and value. The header is inserted at the end of an iteration, which guarantee that we do not insert name only in the list in case of an error on value decoding. This also helps the function readability by centralizing the LIST insert operation. The return value of the decoding function is also changed. Now on success the number of headers inserted in the input list is returned. This change as no impact as success value is not used by the caller. This is mainly done to have a behavior similar to hpack decoding function.	2022-06-15 15:05:22 +02:00
Amaury Denoyelle	60ef19f137	BUG/MINOR: h3/qpack: deal with too many headers ensures that we never insert too many entries in a headers input list. On the decoding side, a new error QPACK_ERR_TOO_LARGE is reported in this case. This prevents crash if headers number on a H3 request or response is superior to tune.http.maxhdr config value. Previously, a crash would occur in QPACK decoding function. Note that the process still crashes later with ABORT_NOW() because error reporting on frame parsing is not implemented for now. It should be treated with a RESET_STREAM frame in most cases. This can be backported up to 2.6.	2022-06-15 15:05:08 +02:00
Amaury Denoyelle	28d3c2489f	MINOR: qpack: add ABORT_NOW on unimplemented decoding Post-base indices is not supported at the moment for decoding. This should never be encountered as it is only used with a dynamic table. However, haproxy deactivates support for the dynamic table via its SETTINGS. Use ABORT_NOW() if this situation happens anyway. This should help debugging instead of silently failed without error reporting.	2022-06-15 14:56:05 +02:00
Amaury Denoyelle	4bcaf69dca	BUG/MINOR: qpack: support header litteral name decoding Complete QPACK decoding with full implementation of litteral field name with litteral value representation. This change is mandatory to support decoding of headers name not present in the QPACK static table. Previously, these headers were silently ignored and not transferred on the backend request. QPACK decoding should now be sufficient to deal with all real situations. Only post-base indices representation are not handled but this should not cause a problem as they are only used for the dynamic table whose size is null as announced by the haproxy implementation. The direct impact of this change is that it should now be possible to use complex webapp through a QUIC frontend. This must be backported up to 2.6.	2022-06-15 14:54:51 +02:00
Amaury Denoyelle	53eef46b88	MINOR: qpack: reduce dependencies on other modules Clean up QPACK decoder API by removing dependencies on ncbuf and MUX-QUIC. This reduces includes statements. It will also help to implement a standalone QPACK decoder.	2022-06-15 11:20:48 +02:00
Amaury Denoyelle	c5d31ed8be	MINOR: qpack: add comments and remove a useless trace Add comments on the decoding function to facilitate code analysis. Also remove the qpack_debug_hexdump() which prints the whole left buffer on each header parsing. With large HEADERS frame payload, QPACK traces are complicated to debug with this statement.	2022-06-15 11:20:42 +02:00
Willy Tarreau	f5aef027ce	OPTIM: task: do not consult shared WQ when we're already full If we've stopped consulting the local wait queue due to too many tasks (max_processed <= 0), there's no point starting to lock the shared WQ, check the first task's expiration date, upgrading the lock just to refrain from doing the work because of the limit. All this does is increase contention on an already contended system. Note that there is still a fairness issue in this WQ dequeuing code. If each thread is busy with expired tasks, no thread will dequeue the global ones. In practice it doesn't make much sense and should quickly resorb, but it could be nice to have an alternating flag indicating where to start from on next call to improve this.	2022-06-14 16:15:15 +02:00
Willy Tarreau	3ccb14d60d	MINOR: thread: get rid of MAX_THREADS_MASK This macro was used both for binding and for lookups. When binding tasks or FDs, using all_threads_mask instead is better as it will later be per group. For lookups, ~0UL always does the job. Thus in practice the macro was already almost not used anymore since the rest of the code could run fine with a constant of all ones there.	2022-06-14 11:18:40 +02:00
Willy Tarreau	e35f03239d	CLEANUP: hlua: check for at least 2 threads on a task In 1.9-dev1, commit `5bc9972ed` ("BUG/MINOR: lua/threads: Make lua's tasks sticky to the current thread") to detect unconfigured Lua tasks that could run on any thread, by comparing their thread mask with MAX_THREADS_MASK. The proper way to do it is to check for at least 2 threads in their mask in fact. This is more reliable and allows to get rid of MAX_THREADS_MASK there.	2022-06-14 11:00:46 +02:00
Willy Tarreau	1a85a958dd	MINOR: tinfo: remove the global thread ID bit (tid_bit) Each thread has its own local thread id and its own global thread id, in addition to the masks corresponding to each. Once the global thread ID can go beyond 64 it will not be possible to have a global thread Id bit anymore, so better start to remove it and use only the local one from the struct thread_info.	2022-06-14 10:44:38 +02:00
Willy Tarreau	8716875ea4	CLEANUP: quic: use task_new_on() for single-threaded tasks This simply replaces a call to task_new(1<<thr) with task_new_on(thr) so that we can later isolate the changes required to add more thread group stuff.	2022-06-14 10:38:03 +02:00
Willy Tarreau	680ed5f28b	MINOR: task: move profiling bit to per-thread Instead of having a global mask of all the profiled threads, let's have one flag per thread in each thread's flags. They are never accessed more than one at a time an are better located inside the threads' contexts for both performance and scalability.	2022-06-14 10:38:03 +02:00
Amaury Denoyelle	040955fb39	BUG/MEDIUM: mux-quic: fix segfault on flow-control frame cleanup LIST_ELEM macro was incorrectly used in the loop when purging flow-control frames from qcc.lfctl.frms on MUX release. This caused a segfault in qc_release() due to an invalid quic_frame pointer instance. The occurence of this bug seems fairly rare. To happen, some flow-control frames must have been allocated but not yet sent just as the MUX release is triggered. I did not find a reproducer scenario. Instead, I artificially triggered it by inserting a quic_frame in qcc.lfctl.frms just before purging it in qc_release() using the following snippet. struct quic_frame *frm; frm = pool_zalloc(pool_head_quic_frame); LIST_INIT(&frm->reflist); frm->type = QUIC_FT_MAX_DATA; frm->max_data.max_data = 0; LIST_APPEND(&qcc->lfctl.frms, &frm->list); This should fix github issue #1747. This must be backported up to 2.6.	2022-06-13 14:43:00 +02:00
Christopher Faulet	4167e05002	BUG/MEDIUM: cli: Notify cli applet won't consume data during request processing The CLI applet process one request after another. Thus, when several requests are pipelined, it is important to notify it won't consume remaining outgoing data while it is processing a request. Otherwise, the applet may be woken up in loop. For instance, it may happen with the HTTP client while we are waiting for the server response if a shutr is received. This patch must be backported in all supported versions after an observation period. But a massive refactoring was performed in 2.6. So, for the 2.5 and below, the patch will have to be adapted. Note also that, AFAIK, the bug can only be triggered by the HTTP client for now.	2022-06-13 14:33:30 +02:00
Christopher Faulet	04f03e15c3	BUG/MEDIUM: stconn: Don't wakeup applet for send if it won't consume data in .chk_snd applet callback function, we must not wake up an applet if SE_FL_WONT_CONSUME flag is set. Indeed, if an applet explicitly specify it will not consume any outgoing data, it is useless to wake it up when more data are sent. Note the applet may still be woken up for another reason. In this case SE_FL_WONT_CONSUME flag will be removed. It is the applet responsibility to set it again, if necessary. This patch must be backported to 2.6 after an observation period. On earlier versions, the bug probably exists too. However, because a massive refactoring was performed in 2.6, the patch will have to be adapted and carefully reviewed/tested if it is backported..	2022-06-13 14:26:13 +02:00
Christopher Faulet	e4b4019280	CLEANUP: check: Remove useless tests on check's stream-connector Since the conn-stream refactoring, from the time the health-check is in progress, its stream-connector is always defined. So, some tests on it are useless and can be removed. This patch should fix the issue #1739.	2022-06-13 08:04:10 +02:00
Christopher Faulet	d1c6cfe45a	BUG/MINOR: tcp-rules: Make action call final on read error and delay expiration When a TCP content ruleset is evaluated, we stop waiting for more data if the inspect-delay is reached, if there is a read error or if we know no more data will be received. This last point is only valid for ACLs. An action may decide to yield for another reason. For instance, in the SPOE, the "send-spoe-group" action yields while the agent response is not received. Thus, now, an action call is final only when the inspect-delay is reached or if there is a read error. But it is possible for an action to yield if the buffer is full or if CF_EOI flag is set. This patch could be backported to all supported versions.	2022-06-13 08:04:10 +02:00
Amaury Denoyelle	43c090c1ed	BUG/MINOR: mux-quic: fix memleak on frames rejected by transport When the MUX transfers a big amount of data to the client, the transport layer may reject some of them because of the congestion controller limit. Frames built by the MUX are thus dropped, even if streams transferred data are kept in buffers for future new frames. Thus, the MUX is required to free rejected frames. This fixes a memory leak which may grow with important data transfers. It should be backported to 2.6 after it has been tested and validated.	2022-06-10 17:59:06 +02:00
Amaury Denoyelle	78fa559679	MINOR: mux-quic: complete BUG_ON on TX flow-control enforcing TX flow-control enforcing is not straightforward : it requires the usage of several counters at stream and connection level, in part due to the difficult sending API between MUX and quic-conn layers. To strengthen this part and ensures it behaves as expected, some existing BUG_ON statements were adjusted and new one were added. This should help to catch errors as early as possible, as in the case with github issue #1738.	2022-06-10 17:40:42 +02:00
Amaury Denoyelle	b9e0640405	BUG/MEDIUM: mux-quic: fix flow control connection Tx level The flow control enforced at connection level is incorrectly calculated. There is a risk of exceeding the limit. In most cases, this results in a segfault produced by a BUG_ON which is here to catch this kind of error. If not compiled with DEBUG_STRICT, this should generate a connection closed by the client due to the flow control overflow. The problem is encountered when transfered payload is big enough to fill the transport congestion window. In this case, some data are rejected by the transport layer and kept by the MUX to be reemitted later. However, these preserved data are not counted on the connection flow control when resubmitted, which gradually amplify the gap between expected and real consumed flow control. To fix this, handle the flow-control at the connection level in the same way as the stream level. A new field qcc.tx.offsets is incremented as soon as data are transfered between stream TX buffers. The field qcc.tx.sent_offsets is preserved to count bytes handled by the transport layer and stop the MUX transfer if limit is reached. As already stated, this bug can occur during transfers with enough emitted data, over multiple streams. When using a single stream, the flow control at the stream level hides it. The BUG_ON crash is reproduced systematically with quiche client : $ quiche-client --no-verify --http-version HTTP/3 -n 10000 https://127.0.0.1:20443/10K This must be backported up to 2.6 when confirmed to work as expected. This should fix github issue #1738.	2022-06-10 17:30:41 +02:00
Willy Tarreau	c6b7a97e54	BUG/MINOR: cli/stats: add missing trailing LF after "show info json" This is the continuation of commit `5a0e7ca5d` ("BUG/MINOR: cli/stats: add missing trailing LF after JSON outputs"). There's also a "show info json" command which was also missing the trailing LF. It's constructed exactly like the "show stat json", in that it dumps a series of fields without any LF. The difference however is that for the stats output, everything was enclosed in an array which required an LF after the closing bracket, while here there's no such array so we have to emit the LF after the loop. That makes the two functions a bit inconsistent, it's quite annoying, but making them better would require sending one LF per line in the stats output, which is not particularly interesting. Given that it took 5+ years to spot that this code wasn't working as expected it doesn't seem worth investing much time trying to refactor it to make it look cleaner at the risk of breaking other obscure parts.	2022-06-10 15:12:21 +02:00
Willy Tarreau	9b46fb4cca	BUG/MINOR: server: do not enable DNS resolution on disabled proxies Leonhard Wimmer reported an interesting bug in github issue #1742. Servers in disabled proxies that are configured for resolution are still subscribed to DNS resolutions, but the LB algos are not initialized at all since the proxy is disabled, so when the server state changes, attempts to update its status cause a crash when the server's weight is recalculated via a divide by the proxy's total weight which is zero. This should be backported to all versions. Beware that before 2.5 or so, there's no PR_FL_DISABLED flag, instead px->disabled should be used (2.3-2.4) or PR_STSTOPPED for older versions. Thanks to Leonhard for his report and quick test!	2022-06-10 11:17:27 +02:00
Willy Tarreau	5a0e7ca5d0	BUG/MINOR: cli/stats: add missing trailing LF after JSON outputs Patrick Hemmer reported that we have a bug in the CLI commands "show stat json" and "show schema json" in that they forget the trailing LF that's required to mark the end of the response. This has been the case since the introduction of the feature in 1.8-dev1 by commit `6f6bb380e` ("MEDIUM: stats: Add show json schema"), so this fix may be backported to all versions.	2022-06-10 09:23:44 +02:00
Amaury Denoyelle	3a2fcfd58d	BUG/MEDIUM: h3: fix SETTINGS parsing Function used to parse SETTINGS frame is incorrect as it does not stop at the frame length but continue to parse beyond it. In most cases, it will result in a connection closed with error H3_FRAME_ERROR. This bug can be reproduced with clients that sent more than just a SETTINGS frame on the H3 control stream. This is notably the case with aioquic which emit a MAX_PUSH_ID after SETTINGS. This bug has been introduced in the current dev release, by the following patch `62eef85961` MINOR: mux-quic: simplify decode_qcs API thus, it does not need to be backported.	2022-06-09 14:34:43 +02:00
Glenn Strauss	0012f899dd	OPTIM: mux-h2: increase h2_settings_initial_window_size default to 64k This changes the default from RFC 7540's default 65535 (64k-1) to avoid avoid some degenerative WINDOW_UPDATE behaviors in the wild observed with clients using 65536 as their buffer size, and have to complete each block with a 1-byte frame, which with some servers tend to degenerate in 1-byte WU causing more 1-byte frames to be sent until the transfer almost only uses 1-byte frames. More details here: https://github.com/nghttp2/nghttp2/issues/1722 As mentioned in previous commit (MEDIUM: mux-h2: try to coalesce outgoing WINDOW_UPDATE frames) the issue could not be reproduced with haproxy but individual WU frames are sent so theoretically nothing prevents this from happening. As such it should be backported as a workaround for already deployed clients after watching for any possible side effect with rare clients. As an added benefit, uploads from curl now use less DATA frames (all are 16384 now). Note that the previous patch alone is sufficient to stop the issue with curl in case this one would need to be reverted. [wt: edited commit messaged, updated doc]	2022-06-09 09:28:21 +02:00
Willy Tarreau	617592c9eb	MEDIUM: mux-h2: try to coalesce outgoing WINDOW_UPDATE frames Glenn Strauss from Lighttpd reported a corner case affecting curl+lighttpd that causes some uploads to degenerate to extremely suboptimal conditions under certain circumstances, and noted that many other implementations were possibly not safe against this degradation. Glenn's detailed analysis is available here: https://github.com/nghttp2/nghttp2/issues/1722 In short, curl uses a 65536 bytes buffer and the default stream window is 65535, with 16384 bytes per frame. Curl will then send 3 frames of 16384 bytes followed by one of 16383, will wait for a window update to send the last byte before recycling the buffer to read the next 64kB. On each round like this, one extra single-byte frame will be sent, and if ACKs for these single-byte frames are not aggregated, this will only allow the client to send one extra byte at a time. At some point it is possible (at least Glenn observed it) to have mostly 1-byte frames in the transfer, resulting in huge CPU usage and a long transfer. It was not possible to reproduce this with haproxy, even when playing with frame sizes, buffer sizes nor window sizes. One reason seems to be that we're using the same buffer size for the connection and the stream and that the frame headers prevent the filling of the window from happening on the same boundaries as on the sender. However it does occasionally happen to see up to two 1-byte data frames in a row, indicating that there's definitely room for improvement. The WINDOW_UPDATE frames for the connection are sent at the end of the demuxing, but the ones for the streams are currently sent immediately after a DATA frame is processed, mostly for convenience. But we don't need to proceed like this, we already have the counter of unacked bytes in rcvd_s, so we can simply use that to decide when to send an ACK. It must just be done before processing a new frame. The benefit is that contiguous frames for the same stream will now only produce a single WU, like for the connection. On complicated tests involving a client that was limited to 100 Mbps transfers and a dummy Lua-based payload consumer, it was possible to see the number of stream WU frames being halved for a 100 MB transfer, which is already a nice saving anyway. Glenn proposed a better workaround consisting in increasing the default window size to 65536. This will be done in a separate patch so that both can be studied independently in field and backported as needed. This patch is not much complicated and shold be backportable. It just needs to be tested in development first.	2022-06-09 09:28:21 +02:00
Amaury Denoyelle	1cd43aa194	BUG/MINOR: h3: fix incorrect BUG_ON assert on SETTINGS parsing BUG_ON() assertion to check for incomplete SETTINGS frame is incorrect. It should check if frame length is greater, not smaller, than current buffer data. Anyway, this BUG_ON() is useless as h3_decode_qcs() prevents parsing of an incomplete frame, except for H3 DATA. Remove it to fix this bug. This bug was introduced in the current dev tree by commit commit `62eef85961` MINOR: mux-quic: simplify decode_qcs API Thus it does not need to be backported. This fixes crashes which happen with DEBUG_STRICT=2. Most notably, this is reproducible with clients that emit more than just a SETTINGS frame on the H3 control stream. It can be reproduced with aioquic for example.	2022-06-08 18:26:06 +02:00
Christopher Faulet	58e3501910	BUG/MEDIUM: mailers: Set the object type for check attached to an email alert The health-check attached to an email alert has no type. It is unexpected, and since the 2.6, it is important because we rely on it to know the application type in front of a connection at the stream-connector level. Because the object type is not set, the SE descriptor is not properly initialized, leading to a segfault when a connection to the SMTP server is established. This patch must be backported to 2.6 and may be backported as far as 2.0. However, it is only an issue for the 2.6 and upper.	2022-06-08 15:28:38 +02:00
Christopher Faulet	4f1825c5db	BUG/MINOR: checks: Properly handle email alerts in trace messages There is no server for email alerts. So the trace messages must be adapted to handle this case. Information related to the server are now skipped for email alerts and "[EMAIL]" prefix is used. This patch must be backported as far as 2.4.	2022-06-08 15:28:38 +02:00
Christopher Faulet	e3b2574796	BUG/MINOR: trace: Test server existence for health-checks to get proxy Email alerts are based on health-checks but with no server. Thus, in __trace() function, responsible to write a trace message, we must be prepared to have no server and thus no proxy. This patch must be backported as far as 2.4.	2022-06-08 15:28:38 +02:00
Benoit DOLEZ	69e3f05b15	BUILD: quic: fix anonymous union for gcc-4.4 Building QUIC with gcc-4.4 on el6 shows this error: src/xprt_quic.c: In function 'qc_release_lost_pkts': src/xprt_quic.c:1905: error: unknown field 'loss' specified in initializer compilation terminated due to -Wfatal-errors. make: * [src/xprt_quic.o] Error 1 make: * Waiting for unfinished jobs.... Initializing an anonymous form of union like : struct quic_cc_event ev = { (...) .loss.time_sent = newest_lost->time_sent, (...) }; generates an error with gcc-4.4 but not when initializing the fields outside of the declaration.	2022-06-08 11:24:36 +02:00
Amaury Denoyelle	dca4c53a95	BUG/MINOR: h3: fix return value on decode_qcs on error Convert return code to -1 when an error has been detected. This is required since the previous API change on return value from the patch : `1f21ebdd76` MINOR: mux-quic/h3: adjust demuxing function return values Without this, QUIC MUX won't consider the call as an error and will try to remove one byte from the buffer. This may cause a BUG_ON failure if the buffer is empty at this stage. This bug was introduced in the current dev tree. Does not need to be backported.	2022-06-07 18:26:46 +02:00
Amaury Denoyelle	1f21ebdd76	MINOR: mux-quic/h3: adjust demuxing function return values Clean the API used by decode_qcs() and transcoder internal functions. Parsing functions now returns a ssize_t which represents the number of consumed bytes or a negative error code. The total consumed bytes is returned via decode_qcs(). The API is now unified and cleaner. The MUX can thus simply use the return value of decode_qcs() instead of substracting the data bytes in the buffer before and after the call. Transcoders functions are not anymore obliged to remove consumed bytes from the buffer which was not obvious.	2022-06-07 18:15:47 +02:00
Amaury Denoyelle	62eef85961	MINOR: mux-quic: simplify decode_qcs API Slightly modify decode_qcs function used by transcoders. The MUX now gives a buffer instance on which each transcoder is free to work on it. At the return of the function, the MUX removes consume data from its own buffer. This reduces the number of invocation to qcs_consume at the end of a full demuxing process. The API is also cleaner with the transcoders not responsible of calling it with the risk of having the input buffer freed if empty.	2022-06-07 18:15:47 +02:00
Amaury Denoyelle	c0156790e6	MINOR: h3: add h3c pointer into h3s instance As a mirror to qcc/qcs types, add a h3c pointer into h3s struct. This should help to clean up H3 code and avoid to use qcs.qcc.ctx to retrieve the h3c instance.	2022-06-07 18:13:11 +02:00
Amaury Denoyelle	16f3da4624	MINOR: connection: support HTTP/3.0 for smp_*_http_major fetch smp_fc_http_major may be used to return the http version as an integer used on the frontend or backend side. Previously, the handler only checked for version 2 or 1 as a fallback. Extend it to support version 3 with the QUIC mux.	2022-06-07 12:04:12 +02:00
Christopher Faulet	1f90f33b69	BUG/MINOR: ssl_ckch: Fix another possible uninitialized value Commit `d6c66f06a` ("MINOR: ssl_ckch: Remove service context for "set ssl crl-file" command") introduced a regression leading to a build error because of a possible uninitialized value. It is now fixed. This patch must be backported as far as 2.5.	2022-06-03 16:49:53 +02:00
Christopher Faulet	ea2c8c6ba7	BUILD: ssl_ckch: Fix build error about a possible uninitialized value A build error is reported about the path variable in the switch statement on the commit type, in cli_io_handler_commit_cafile_crlfile() function. The enum contains only 2 values, but a default clause has been added to return an error to make GCC happy. This patch must be backported as far as 2.5.	2022-06-03 16:41:11 +02:00
Christopher Faulet	88041b35c3	BUG/MINOR: ssl_ckch: Fix possible uninitialized value in show_crlfile I/O handler Commit `9a99e5478` ("BUG/MINOR: ssl_ckch: Dump CRL transaction only once if show command yield") introduced a regression leading to a build error because of a possible uninitialized value. It is now fixed. This patch must be backported as far as 2.5.	2022-06-03 16:28:11 +02:00
Christopher Faulet	677cb4fa91	BUG/MINOR: ssl_ckch: Fix possible uninitialized value in show_cafile I/O handler Commit `5a2154bf7` ("BUG/MINOR: ssl_ckch: Dump CA transaction only once if show command yield") introduced a regression leading to a build error because of a possible uninitialized value. It is now fixed. This patch must be backported as far as 2.5.	2022-06-03 16:27:49 +02:00
Christopher Faulet	d1d2e4dfe5	BUG/MINOR: ssl_ckch: Fix possible uninitialized value in show_cert I/O handler Commit `3e94f5d4b` ("BUG/MINOR: ssl_ckch: Dump cert transaction only once if show command yield") introduced a regression leading to a build error because of a possible uninitialized value. It is now fixed. This patch must be backported as far as 2.2.	2022-06-03 16:27:41 +02:00
Christopher Faulet	d6c66f06ac	MINOR: ssl_ckch: Remove service context for "set ssl crl-file" command This command does not have I/O handle function. All is done in the command parsing function. So there is no reason to have dedicated context.	2022-06-03 12:12:04 +02:00
Christopher Faulet	132c595673	MINOR: ssl_ckch: Remove service context for "set ssl ca-file" command This command does not have I/O handle function. All is done in the command parsing function. So there is no reason to have dedicated context.	2022-06-03 12:12:04 +02:00
Christopher Faulet	24a20b9808	MINOR: ssl_ckch: Remove service context for "set ssl cert" command This command does not have I/O handle function. All is done in the command parsing function. So there is no reason to have dedicated context.	2022-06-03 12:12:04 +02:00
Christopher Faulet	6af2fc6a3f	MINOR: ssl_ckch: Simplify structure used to commit changes on CA/CRL entries The same type is used for CA and CRL entries. So, in commit_cert_ctx structure, there is no reason to have different fields for the CA and CRL entries.	2022-06-03 12:12:04 +02:00
Christopher Faulet	dd0c4834ef	CLEANUP: ssl_ckch: Remove unused field in commit_cacrlfile_ctx structure .next_ckchi field is not used by functions responsible to commit changes on CA/CRL entries. It can be removed.	2022-06-03 12:12:04 +02:00
Christopher Faulet	f814c4aa98	BUG/MINOR: ssl_ckch: Init right field when parsing "commit ssl crl-file" cmd .next_ckchi_link field must be initialized to NULL instead of .next_ckchi in cli_parse_commit_crlfile() function. Only '.nex_ckchi_link' is used in the I/O handler. This patch must be backported as far as 2.5 with some adaptations for the 2.5.	2022-06-03 12:12:04 +02:00
Christopher Faulet	3e94f5d4b6	BUG/MINOR: ssl_ckch: Dump cert transaction only once if show command yield When loaded SSL certificates are displayed via "show ssl cert" command, the in-progess transaction, if any, is also displayed. However, if the command yield, the transaction is re-displayed again and again. To fix the issue, old_ckchs field is used to remember the transaction was already displayed. This patch must be backported as far as 2.2.	2022-06-03 11:20:41 +02:00
Christopher Faulet	5a2154bf7c	BUG/MINOR: ssl_ckch: Dump CA transaction only once if show command yield When loaded CA files are displayed via "show ssl ca-file" command, the in-progress transaction, if any, is also displayed. However, if the command yield, the transaction is re-displayed again and again. To fix the issue, old_cafile_entry field is used to remember the transaction was already displayed. This patch must be backported as far as 2.5.	2022-06-03 11:20:38 +02:00
Christopher Faulet	9a99e54787	BUG/MINOR: ssl_ckch: Dump CRL transaction only once if show command yield When loaded CRL files are displayed via "show ssl crl-file" command, the in-progess transaction, if any, is also displayed. However, if the command yield, the transaction is re-displayed again and again. To fix the issue, old_crlfile_entry field is used to remember the transaction was already displayed. This patch must be backported as far as 2.5.	2022-06-03 11:20:34 +02:00
Christopher Faulet	51095ee236	BUG/MINOR: ssl_ckch: Use right type for old entry in show_crlfile_ctx Because of a typo (I guess), an unknown type is used for the old entry in show_crlfile_ctx structure. Because this field is unused, there is no compilation error. But it must be a cafile_entry and not a crlfile_entry. Note this field is not used for now, but it will be used. This patch must be backported to 2.6.	2022-06-03 11:20:16 +02:00
Christopher Faulet	ddc8e1cf8b	MINOR: ssl_ckch: Simplify I/O handler to commit changes on CA/CRL entry Simplify cli_io_handler_commit_cafile_crlfile() handler function by retrieving old and new entries at the beginning. In addition the path is also retrieved at this stage. This removes several switch statements. Note that the ctx was already validated by the corresponding parsing function. Thus there is no reason to test the pointers. While it is not a bug, this patch may help to fix issue #1731.	2022-06-03 09:21:47 +02:00
Christopher Faulet	14df913400	CLEANUP: ssl_ckch: Use corresponding enum for commit_cacrlfile_ctx.cafile_type There is an enum to determine the entry entry type when changes are committed on a CA/CRL entry. So use it in the service context instead of an integer. This patch may help to fix issue #1731.	2022-06-03 09:21:47 +02:00
Tim Duesterhus	9fb57e8c17	CLEANUP: Re-apply xalloc_size.cocci (2) This reapplies the xalloc_size.cocci patch across the whole `src/` tree. see `16cc16dd82` see `63ee0e4c01`	2022-06-02 14:12:18 +02:00
Christopher Faulet	d649b57519	MEDIUM: http-ana: Always report rewrite failures as PRXCOND in logs Rewrite failures in http rules are reported as proxy errors (PRXCOND) in logs. However, other rewrite errors are reported as internal errors. For instance, it happens when we fail to add X-Forwarded-For header. It is not consistent and it is confusing. So now, all rewite failures are reported as proxy errors. This patch may be backported if necessary.	2022-06-02 12:21:32 +02:00
Christopher Faulet	89f2626c19	MEDIUM: httpclient: Don't close CLI applet at the end of a response There is no reason to close the CLI applet when the whole response was dumped. This prevent anyone to use the CLI in interactive mode.	2022-06-01 17:20:57 +02:00
Christopher Faulet	0158bb23d7	BUG/MEDIUM: httpclient: Rework CLI I/O handler to handle full buffer cases 'httpclient' command does not properly handle full buffer cases. When the response buffer is full, we exit to retry later. However, the context flags are updated. It means when this happens, we may loose a part of the response. So now, flags are preserved when we fail to push data into the response buffer. In addition, instead of dumping one part per call, we now try to dump as much data as possible. Finally, when there is no more data, because everything was dumped or because we are waiting for more data from the HTTP client, the applet is updated accordingly by calling applet_have_no_more_data(). Otherwise, when some data are blocked, applet_putchk() already takes care to update the SE flags. So, it is useless to call sc_need_room(). This patch should fix the issue #1723. It must be backported as far as 2.5. But a massive refactoring was performed in 2.6. So, for the 2.5 and below, the patch will have to be adapted.	2022-06-01 17:20:57 +02:00
Christopher Faulet	18de6f2880	BUG/MEDIUM: httpclient: Don't remove HTX header blocks before duplicating them Commit `534645d6` ("BUG/MEDIUM: httpclient: Fix loop consuming HTX blocks from the response channel") introduced a regression. When the response is consumed, The HTX header blocks are removed before duplicating them. Thus, the first header block is always lost. This patch must be backported as far as 2.5.	2022-06-01 17:20:57 +02:00
Christopher Faulet	c642d7c131	BUG/MEDIUM: ssl/crt-list: Rework 'add ssl crt-list' to handle full buffer cases 'add ssl crt-list' command is also concerned. This patch is similar to the previous ones. Full buffer cases when we try to push the reply are not properly handled. To fix the issue, the functions responsible to add a crt-list entry were reworked. First, the error message is now part of the service context. This way, if we cannot push the error message in the reponse buffer, we may retry later. To do so, a dedicated state was created (ADDCRT_ST_ERROR,). Then, the success message is also handled in a dedicated state (ADDCRT_ST_SUCCESS). This way we are able to retry to push it if necessary. Finally, the dot displayed for each new instance is now immediatly pushed in the response buffer, and before the update. This way, we are able to retry too if necessary. This patch should fix the issue #1724. It must be backported as far as 2.2. But a massive refactoring was performed in 2.6. So, for the 2.5 and below, the patch will have to be adapted.	2022-06-01 17:20:57 +02:00
Christopher Faulet	e9c3bd1395	BUG/MEDIUM: ssl_ckch: Rework 'commit ssl ca-file' to handle full buffer cases 'commit ssl crl-file' command is also concerned. This patch is similar to the previous one. Full buffer cases when we try to push the reply are not properly handled. To fix the issue, the functions responsible to commit CA or CRL entry changes were reworked. First, the error message is now part of the service context. This way, if we cannot push the error message in the reponse buffer, we may retry later. To do so, a dedicated state was created (CACRL_ST_ERROR). Then, the success message is also handled in a dedicated state (CACRL_ST_SUCCESS). This way we are able to retry to push it if necessary. Finally, the dot displayed for each updated CKCH instance is now immediatly pushed in the response buffer, and before the update. This way, we are able to retry too if necessary. This patch should fix the issue #1722. It must be backported as far as 2.5. But a massive refactoring was performed in 2.6. So, for the 2.5, the patch will have to be adapted.	2022-06-01 17:20:57 +02:00
Christopher Faulet	9d56e248a6	BUG/MEDIUM: ssl_ckch: Rework 'commit ssl cert' to handle full buffer cases When changes on a certificate are commited, a trash buffer is used to create the response. Once done, the message is copied in the response buffer. However, if the buffer is full, there is no way to retry and the message is lost. The same issue may happen with the error message. It is a design issue of cli_io_handler_commit_cert() function. To fix it, the function was reworked. First, the error message is now part of the service context. This way, if we cannot push the error message in the reponse buffer, we may retry later. To do so, a dedicated state was created (CERT_ST_ERROR). Then, the success message is also handled in a dedicated state (CERT_ST_SUCCESS). This way we are able to retry to push it if necessary. Finally, the dot displayed for each updated CKCH instance is now immediatly pushed in the response buffer, and before the update. This way, we are able to retry too if necessary. This patch should fix the issue #1725. It must be backported as far as 2.2. But massive refactoring was performed in 2.6. So, for the 2.5 and below, the patch must be adapted.	2022-06-01 17:20:57 +02:00
Christopher Faulet	1e00c7e8f4	BUG/MINOR: ssl_ckch: Don't duplicate path when replacing a CA/CRL entry When a CA or CRL entry is replaced (via 'set ssl ca-file' or 'set ssl crl-file' commands), the path is duplicated and used to identify the ongoing transaction. However, if the same command is repeated, the path is still duplicated but the transaction is not changed and the duplicated path is not released. Thus there is a memory leak. By reviewing the code, it appears there is no reason to duplicate the path. It is always the filename path of the old entry. So, a reference on it is now used. This simplifies the code and this fixes the memory leak. This patch must be backported as far as 2.5.	2022-06-01 16:28:15 +02:00
Christopher Faulet	e2ef4dd3c5	BUG/MINOR: ssl_ckch: Don't duplicate path when replacing a cert entry When a certificate entry is replaced (via 'set ssl cert' command), the path is duplicated and used to identify the ongoing transaction. However, if the same command is repeated, the path is still duplicated but the transaction is not changed and the duplicated path is not released. Thus there is a memory leak. By reviewing the code, it appears there is no reason to duplicate the path. It is always the path of the old entry. So, a reference on it is now used. This simplifies the code and this fixes the memory leak. This patch must be backported as far as 2.2.	2022-06-01 16:28:15 +02:00
Christopher Faulet	1f08fa46fb	BUG/MEDIUM: ssl_ckch: Don't delete CA/CRL entry if it is being modified When a CA or a CRL entry is being modified, we must take care to no delete it because the corresponding ongoing transaction still references it. If we do so, it leads to a null-deref and a crash may be exeperienced if changes are commited. This patch must be backported as far as 2.5.	2022-06-01 16:28:15 +02:00
Christopher Faulet	926fefca8d	BUG/MEDIUM: ssl_ckch: Don't delete a cert entry if it is being modified When a certificate entry is being modified, we must take care to no delete it because the corresponding ongoing transaction still references it. If we do so, it leads to a null-deref and a crash may be exeperienced if changes are commited. This patch must be backported as far as 2.2.	2022-06-01 16:28:15 +02:00
Christopher Faulet	4329dcc2fc	BUG/MINOR: ssl_ckch: Free error msg if commit changes on a CA/CRL entry fails On the CLI, If we fail to commit changes on a CA or a CRL entry, an error message is returned. This error must be released. This patch must be backported as far as 2.4.	2022-06-01 16:28:15 +02:00
Christopher Faulet	01a09e24ad	BUG/MINOR: ssl_ckch: Free error msg if commit changes on a cert entry fails On the CLI, If we fail to commit changes on a certificate entry, an error message is returned. This error must be released. This patch must be backported as far as 2.2.	2022-06-01 16:28:15 +02:00
Amaury Denoyelle	5869cb669e	BUG/MINOR: qpack: do not consider empty enc/dec stream as error When parsing QPACK encoder/decoder streams, h3_decode_qcs() displays an error trace if they are empty. Change the return code used in QPACK code to avoid this trace. To uniformize with MUX/H3 code, 0 is now used to indicate success. Beyond this spurious error trace, this bug has no impact.	2022-05-31 15:35:06 +02:00
Amaury Denoyelle	9f17a5aa8a	CLEANUP: quic: remove useless check on local UNI stream reception The MUX now provides a single API for both uni and bidirectional streams. It is responsible to reject reception on a local unidirectional stream with the error STREAM_STATE_ERROR. This is already implemented in qcc_recv(). As such, remove this duplicated check from xprt_quic.c.	2022-05-31 15:21:13 +02:00
Frédéric Lécaille	fdc1b96357	BUG/MINOR: quic: Fix QUIC_EV_CONN_PRSAFRM event traces This is a quic_frame struct pointer which must be passed as parameter to TRACE_PROTO() for such an event.	2022-05-31 14:46:02 +02:00
Amaury Denoyelle	417c7c03f4	BUG/MEDIUM: h3: fix H3_EXCESSIVE_LOAD when receiving H3 frame header only The H3 frame demuxing code is incorrect when receiving a STREAM frame which contains only a new H3 frame header without its payload. In this case, the check on frames bigger than the buffer size is incorrect. This is because the buffer has been freed via qcs_consume()/qc_free_ncbuf() as it was emptied after H3 frame header parsing. This causes the connection to be incorrectly closed with H3_EXCESSIVE_LOAD error. This bug was reproduced with xquic client on the interop and with the command-line invocation : $ ./interop_client -l d -k $SSLKEYLOGFILE -a <addr> -p <port> -D /tmp \ -A h3 -U https://<addr>:<port>/hello_world.txt Note also that h3_is_frame_valid() invocation has been moved before the new buffer size check. This ensures that first we check the frame validity before returning from the function. It's also better positionned as this is only needed when a new H3 frame header has been parsed.	2022-05-31 14:45:51 +02:00
Amaury Denoyelle	88d5dd1a6f	BUG/MINOR: h3: fix frame demuxing The H3 demuxing code was not fully correct. After parsing the H3 frame header, the check between frame length and buffer data is wrong as we compare a copy of the buffer made before the H3 header removal. Fix this by improving the H3 demuxing code API. h3_decode_frm_header() now uses a ncbuf instance, this prevents an unnecessary cast ncbuf/buffer in h3_decode_qcs() which resolves this error. This bug was not triggered at this moment. Its impact should be really limited.	2022-05-31 14:43:14 +02:00
Amaury Denoyelle	1194db24bc	MINOR: ncbuf: adjust ncb_data with NCBUF_NULL Replace ncb_blk_is_null() by ncb_is_null() as a prelude to ncb_data(). The result is the same : the function will return 0 if the buffer is uninitialized. However, it is clearer to directly call ncb_is_null() to reflect this. There is no functional change with this commit.	2022-05-31 14:31:48 +02:00
Willy Tarreau	824c8c5999	BUG/MINOR: peers: detect and warn on init_addr/resolvers/check/agent-check Some server keywords are currently silently ignored in the peers section, which is not good because it wastes time on user-side, trying to make something work while it cannot by design. With this patch we at least report a few of them (the most common ones), which are init_addr, resolvers, check, agent-check. Others might follow. This may be backported to 2.5 to encourage some cleaning of bogus configs.	2022-05-31 09:42:44 +02:00
Willy Tarreau	245721b329	MINOR: server: indicate when no address was expected for a server When parsing a peers section, it's particularly difficult to make the difference between the local peer which doesn't have any address, and other peers which need one, and the error messages do not help because with just: peers foo bind :8001 server foo 127.0.0.1:8001 server bar 127.0.0.2:8001 One can get such a confusing message when the local peer is "bar": [peers.cfg:15] : 'server foo/bar' : unknown keyword '127.0.0.1:8001'. It's not clear there why the other peer doesn't trigger an error. With this commit we add a hint in the error message when no address was expected. The error remains quite generic (since deep into the server code) but at least the useer gets a hint about why the keyword wasn't understood: [peers.cfg:15] : 'server foo/bar' : unknown keyword '127.0.0.1:8001'. Hint: no address was expected for this server.	2022-05-31 09:25:34 +02:00
Willy Tarreau	356866acce	BUG/MINOR: peers: set the proxy's name to the peers section name For some poor historical reasons, the name of a peers proxy used to be set to the name of the local peer itself. That causes some confusion when multiple sections are present because the same proxy name appears at multiple places in "show peers", but since 2.5 where parsing errors include the proxy name, a config like this one : peers foo server foobar blah Would report this when the local peer name isn't "foobar": 'server (null)/foobar' : invalid address: 'blah' in 'blah' And this when it is foobar: 'server foobar/foobar' : invalid address: 'blah' in 'blah' This is wrong, confusing and not very practical. This commit addresses all this by using the peers section's name when it's created. This now allows to report messages such as: 'server foo/foobar' : invalid address: 'blah' in 'blah' Which make it clear that the section is called "foo" and the server "foobar". This may be backported to 2.5, though the patch may be simplified if needed, by just adding the change at the output of init_peers_frontend().	2022-05-31 09:10:19 +02:00
Willy Tarreau	50e77b2b85	CLEANUP: peers/cli: make peers_dump_peer() take an appctx instead of an stconn By having the appctx in argument this function wouldn't have experienced the previous bug. Better do that now to avoid proliferation of awkward functions.	2022-05-31 08:55:54 +02:00
Willy Tarreau	fc5059958f	CLEANUP: peers/cli: stop misusing the appctx local variable In the context of a CLI command, it's particularly not welcome to use an "appctx" variable that is not the current one. In addition it was created for use at exactly 6 places in 2 lines. Let's just remove it and stick to peer->appctx which is used elsewhere in the function and is unambiguous.	2022-05-31 08:53:25 +02:00
Willy Tarreau	ccea010104	BUG/MEDIUM: peers/cli: fix "show peers" crash Commit `d0a06d52f` ("CLEANUP: applet: use applet_put*() everywhere possible") replaced most accesses to the conn_stream with simpler accesses to the appctx. Unfortunately, in all the CLI functions using an appctx, one makes an exception where the appctx is not the caller's but the one being inspected! When no peers connection is active, the early exit immediately crashes. No backport is needed.	2022-05-31 08:49:29 +02:00
Amaury Denoyelle	d5581d527c	MINOR: h3: add traces on h3s init/end Add events when h3s instances are created/initialized and released.	2022-05-30 17:37:50 +02:00
Amaury Denoyelle	a717eb7136	MINOR: h3: add traces on frame send Add h3 traces events for several sent frames : SETTINGS, HEADERS and DATA.	2022-05-30 17:36:39 +02:00
Amaury Denoyelle	494512d00f	MINOR: h3: add traces on frame recv Add h3 traces events for several received frames : SETTINGS, HEADERS and DATA.	2022-05-30 17:35:57 +02:00
Amaury Denoyelle	016aa93088	MINOR: h3: define h3 trace module A new 'h3' trace module is introduced. It will be used to centralize events related to HTTP/3 status.	2022-05-30 17:34:51 +02:00
Willy Tarreau	d46b5b94f0	BUILD: htx: use the unchecked version of htx_get_head_blk() where needed stream.c and mux_fcgi.c may cause a warning for a possible NULL deref at -Os, while that is not possible thanks to the previous test. Let's just switch to __htx_get_head_blk() instead.	2022-05-30 16:27:48 +02:00
Amaury Denoyelle	b93399a5e7	BUG/MINOR: h3: do not report bug on unknown method Remove an unneeded BUG_ON statement when find_http_meth() returns HTTP_METH_OTHER. This fix is necessary to support requests with unusual methods with DEBUG_STRICT activated. This was detected when browsing with HTTP/3 over a nextcloud instance which uses PROPFIND method for Webdav.	2022-05-30 14:30:05 +02:00
Amaury Denoyelle	11f5a796c1	BUG/MINOR: qpack: support bigger prefix-integer encoding Prefix-integer encoding function was incomplete. It was not able to deal correctly with value encoded on more than 2 bytes. This maximum value depends on the size of the prefix, but value greater than 254 were all impacted. Most notably, this change is required to support header name/value with sizeable length. Previously, length was incorrectly encoded. The client thus closed the connection with QPACK_DECOMPRESSION_ERROR.	2022-05-30 14:30:05 +02:00
Amaury Denoyelle	5f6de8d77a	BUG/MINOR: qpack: fix buffer API usage on prefix integer encoding Replace bogus call b_data() by b_room() to check if there is enough space left in the buffer before encoding a prefix integer. At this moment, no real scenario was found to trigger a bug related to this change. This is probably because the buffer always contains data (field section line and status code) before calling qpack_encode_prefix_integer() which prevents an occurrence of this bug.	2022-05-30 14:28:46 +02:00
Frédéric Lécaille	e06ca65e8d	MINOR: quic: Do not drop packets with RESET_STREAM frames If the connection client timeout has expired, the mux is released. If the client decides to initiate a new request, we send a STOP_SENDING frame. Then, the client endessly sends a RESET_STREAM frame. At this time, we simulate the fact that we support the RESET_STREAM frame thanks to this ridiculously minimalistic patch.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	4df2fe90c8	MINOR: quic: Send STOP_SENDING frames if mux is released If the connection client timeout has expired, the mux is released. If the client decides to initiate a new request, we do not ack its request. This leads the client to endlessly sent it request. This patch makes a QUIC listener send a STOP_SENDING frame in such a situation.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	6f7607ef1f	MINOR: h3: Add a statistics module for h3 Add ->inc_err_cnt new callback to qcc_app_ops struct which can be called from xprt to increment the application level error code counters. It take the application context as first parameter to be generic and support new QUIC applications to come. Add h3_stats.c module with counters for all the frame types and error codes.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	38dea05ca9	MINOR: quic: Connection TX buffer setting renaming. Rename "tune.quic.conn-buf-limit" to "tune.quic.frontend.conn-tx-buffers.limit" to reflect the stream direction (TX) and the objects (frontends) which are concerned.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	eb79145f01	MINOR: quic_stats: Add transport new counters (lost, stateless reset, drop) Add new counters to count the number of dropped packet upon parsing error, lost sent packets and the number of stateless reset packet sent. Take the oppportunity of this patch to rename CONN_OPENINGS to QUIC_ST_HALF_OPEN_CONN (total number of half open connections) and QUIC_ST_HDSHK_FAILS to QUIC_ST_HDSHK_FAIL.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	91a211fb08	BUG/MINOR: quic: Largest RX packet numbers mixing When we select the next encryption level in qc_treat_rx_pkts() we must reset the local largest_pn variable if we do not want to reuse its previous value for this encryption. This bug could only happend during handshake step and had no visible impact because this variable is only used during the header protection removal step which hopefully supports the packet reordering.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	3ccea6d276	MINOIR: quic_stats: add QUIC connection errors counters Add statistical counters for all the transport level connection errrors.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	aee675746c	MINOR: quic: Clarifications about transport parameters value This is becoming difficult to distinguish the default values for transport parameters which come with the RFC from our implementation default values when not set by configuration (tunable parameters). Add a comment to distinguish them. Prefix these default values by QUIC_TP_DFLT_ to distinguish them from QUIC_DFLT_* value even if there are not numerous. Furthermore ->max_udp_payload_size must be first initialized to QUIC_TP_DFLT_MAX_UDP_PAYLOAD_SIZE especially for received value.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	2674098569	MINOR: quic: Tunable "initial_max_streams_bidi" transport parameter Add tunable "tune.quic.frontend.max_streams_bidi" setting for QUIC frontends to set the "initial_max_streams_bidi" transport parameter. Add some documentation for this new setting.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	1d96d6e024	MINOR: quic: Tunable "max_idle_timeout" transport parameter Add two tunable settings both for backends and frontends "max_idle_timeout" QUIC transport parameter, "tune.quic.frontend.max-idle-timeout" and "tune.quic.backend.max-idle-timeout" respectively. cfg_parse_quic_time() has been implemented to parse a time value thanks to parse_time_err(). It should be reused for any tunable time value to be parsed. Add the documentation for this tunable setting only for frontend.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	c7785b5c26	MINOR: quic: Transport parameters dump Add quic_transport_params_dump() static inline function to do so for a quic_transport_parameters struct as parameter. We use the trace API do dump these transport parameters both after they have been initialized (RX/local) or received (TX/remote).	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	748ece68b8	MINOR: quic: QUIC transport parameters split. Make the transport parameters be standlone as much as possible as it consists only in encoding/decoding data into/from buffers. Reduce the size of xprt_quic.h. Unfortunalety, I think we will have to continue to include <xprt_quic-t.h> to use the trace API into this module.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	57ac3faed7	CLEANUP: quic: No more used handshake output buffer ->obuf quic_conn struct member is no more used.	2022-05-30 09:59:26 +02:00
Frédéric Lécaille	f6954c5c3a	MINOR: quic: Ignore out of packet padding. We do not want to count the out of packet padding as being belonging to an invalid packet, the firt byte of a QUIC packet being never null. Some browsers like firefox proceeds this way to add PADDING frames after an Initial packet and increase the size of their Initial packets.	2022-05-30 09:59:26 +02:00
Christopher Faulet	186367f499	CLEANUP: muxes: Consider stream's sd as defined in .show_fd callback functions In muxes, the stream-endoint descriptor of a stream is always defined. Thus, in .show_fd callback functions, there is no reason to test it. This patch should address the issue #1727. .	2022-05-30 08:45:16 +02:00
Christopher Faulet	560b8da258	CLEANUP: tcpcheck: Remove useless test on the stream-connector in tcpcheck_main Thanks to the recent refactoring, when tcpcheck_main() function is called, the stream-connector of the healthchek is always defined. There is no reason to still test it. This patch should fix the issue #1721.	2022-05-30 08:37:40 +02:00
Willy Tarreau	da59c895b9	CLEANUP: stconn: remove the new unneeded SE_FL_APP_MASK The only two places where it was used was to carefully preserve the SE_FL_WILL_CONSUME flag (since others are irrelevant there and the previous RXBLK* flags moved to the stconn). Now that the flag is cleared by default there's no need to re-created a fresh new one when replacing the descriptor, so we can eliminate that remaining trick.	2022-05-27 19:33:35 +02:00
Willy Tarreau	369d5aa208	CLEANUP: stream: remove unneeded test on appctx during initialization Now that the data consumption from the endpoint is the default setting, we can generalize the pre-clearing of the wont_consume flag, which is no more specific to applets. In practice it's not needed anymore to do it, but since streams might be initiatied from asynchronous applets, these might have blocked their consumption side before creating the stream thus it's safer to preserve the clearing of the flag.	2022-05-27 19:33:35 +02:00
Willy Tarreau	3121928f82	CLEANUP: stconn: rename a few "endp" arguments and variables to "sd" For consistency with the few other places, let's avoid using the confusing "endp" pointer when it designates the descriptor.	2022-05-27 19:33:35 +02:00
Willy Tarreau	9e00da1f60	CLEANUP: mux-pt: rename the "endp" field to "sd" The stream endpoint descriptor that was named "endp" is now called "sd" both in the mux_pt_ctx struct and in the few functions using this.	2022-05-27 19:33:35 +02:00
Willy Tarreau	5aa5e77cad	CLEANUP: mux-fcgi: rename the "endp" field to "sd" The stream endpoint descriptor that was named "endp" is now called "sd" both in the fcgi_strm struct and in the few functions using this. The name was also updated in the "show fd" output.	2022-05-27 19:33:35 +02:00
Willy Tarreau	95acc8b07f	CLEANUP: mux-h2: rename the "endp" field to "sd" The stream endpoint descriptor that was named "endp" is now called "sd" both in the h2s struct and in the few functions using this. The name was also updated in the "show fd" output.	2022-05-27 19:33:35 +02:00
Willy Tarreau	1a0d9acd3b	CLEANUP: mux-h1: rename the "endp" field to "sd" The stream endpoint descriptor that was named "endp" is now called "sd" both in the h1s struct and in the few functions using this. The name was also updated in the "show fd" output.	2022-05-27 19:33:35 +02:00
Willy Tarreau	d7b7e0df9a	CLEANUP: mux-quic: rename the "endp" field to "sd" The stream endpoint descriptor that was named "endp" is now called "sd" both in the qcs struct and in the few functions using this.	2022-05-27 19:33:35 +02:00
Willy Tarreau	e68bc6178a	CLEANUP: stconn: replace a few remaining occurrences of CS in comments or traces A few "CS" desginating stconns were still present in code comments and stream traces. This addresses them.	2022-05-27 19:33:35 +02:00
Willy Tarreau	1d2c79a53c	CLEANUP: obj_type: rename OBJ_TYPE_CS to OBJ_TYPE_SC Let's apply the new name to the type as well.	2022-05-27 19:33:35 +02:00
Willy Tarreau	df1a2fc234	CLEANUP: stream: rename stream_upgrade_from_cs() to stream_upgrade_from_sc() It upgrades the protocol on a stream connector, let's update the name.	2022-05-27 19:33:35 +02:00
Willy Tarreau	c12b321661	CLEANUP: applet: rename appctx_cs() to appctx_sc() It returns a stream connector, not a conn_stream anymore, so let's fix its name.	2022-05-27 19:33:35 +02:00
Willy Tarreau	b52d4d217f	CLEANUP: sslsock: remove only occurrence of local variable "cs" In ssl_action_wait_for_hs() the local variables called "cs" is just a copy of s->scf that's only used once, so it can be removed. In addition the check was removed as well since it's not possible to have a NULL SC on a stream.	2022-05-27 19:33:35 +02:00
Willy Tarreau	0eca539dbd	CLEANUP: sink: rename all occurrences of stconn "cs" to "sc" In the applet, function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	da30490b9c	CLEANUP: peers: rename all occurrences of stconn "cs" to "sc" In the applet, function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	7577d9d99c	CLEANUP: mux-pt: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion. There was also one place in traces where "cs" used to display the stconn, which were turned to "sc".	2022-05-27 19:33:35 +02:00
Willy Tarreau	36c223243f	CLEANUP: mux-h2: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion. There were also 2 places in traces where "cs" used to display the stconn, which were turned to "sc". The "nb_cs" struct field and "h2_has_too_many_cs()" functions were also renamed.	2022-05-27 19:33:35 +02:00
Willy Tarreau	000d63cfd8	CLEANUP: mux-h1: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion. There were also 2 places in traces where "cs" used to display the stconn, which were turned to "sc". h1s_upgrade_cs() and h1s_new_cs() were both renamed to _cs.	2022-05-27 19:33:35 +02:00
Willy Tarreau	c92a6ca476	CLEANUP: mux-fcgi: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion. There were also 3 places in debugging traces where "cs" used to display the stconn, which were turned to "sc" for similar reasons. The number of streams "nb_cs" was turned to "nb_sc".	2022-05-27 19:33:35 +02:00
Willy Tarreau	b89f872947	CLEANUP: http-client: rename all occurrences of stconn "cs" to "sc" In the applet, function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	270a4574a4	CLEANUP: log-forward: rename all occurrences of stconn "cs" to "sc" In the log-forwarding applet, function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	3e7be363c4	CLEANUP: hlua: rename all occurrences of stconn "cs" to "sc" In the TCP and HTTP applets, function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	9002a2e416	CLEANUP: spoe: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables in the SPOE applet called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	d7950ad447	CLEANUP: dns: rename all occurrences of stconn "cs" to "sc" This concerns the DNS client applet. Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	c5ddd9fd80	CLEANUP: cache: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	07ee044762	CLEANUP: applet: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	475e4636bc	CLEANUP: cli: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" in the various keyword handlers.	2022-05-27 19:33:35 +02:00
Willy Tarreau	caff631bc0	CLEANUP: stats: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion. Both the core functions and the ones in the resolvers files were updated.	2022-05-27 19:33:35 +02:00
Willy Tarreau	b49672d21f	CLEANUP: stream: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion. The HTTP analyser and the backend functions were all updated after being reviewed. Function stream_update_both_cs() was renamed to stream_update_both_sc()	2022-05-27 19:33:35 +02:00
Willy Tarreau	3215e731b6	CLEANUP: quic/h3: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion. The "nb_cs" stream-connector counter was renamed to "nb_sc" and qc_attach_cs() was renamed to qc_attach_sc().	2022-05-27 19:33:35 +02:00
Willy Tarreau	0adb281fb0	CLEANUP: stconn: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion. The change is huge (~580 lines), so extreme care was given not to change anything else.	2022-05-27 19:33:35 +02:00
Willy Tarreau	61f5675cb4	CLEANUP: connection: rename all occurrences of stconn "cs" to "sc" Function arguments and local variables called "cs" were renamed to "sc" to avoid future confusion.	2022-05-27 19:33:35 +02:00
Willy Tarreau	bde14ad499	CLEANUP: check: rename all occurrences of stconn "cs" to "sc" The check struct had a "cs" field renamed to "sc", which also required a tiny update to a few functions using it to distinguish a check from a stream (log.c, payload.c, ssl_sample.c, tcp_sample.c, tcpcheck.c, connection.c). Function arguments and local variables called "cs" were renamed to "sc". The presence of one "cs=" in the debugging traces was also turned to "sc=" for consistency.	2022-05-27 19:33:35 +02:00
Willy Tarreau	d137353ae3	CLEANUP: muxes: rename "get_first_cs" to "get_first_sc" This is renamed both in the mux_ops descriptor and the mux functions themselves to accommodate the new type name.	2022-05-27 19:33:35 +02:00
Willy Tarreau	cb086c6de1	REORG: stconn: rename conn_stream.{c,h} to stconn.{c,h} There's no more reason for keepin the code and definitions in conn_stream, let's move all that to stconn. The alphabetical ordering of include files was adjusted.	2022-05-27 19:33:35 +02:00
Willy Tarreau	5edca2f0e1	REORG: rename cs_utils.h to sc_strm.h This file contains all the stream-connector functions that are specific to application layers of type stream. So let's name it accordingly so that it's easier to figure what's located there. The alphabetical ordering of include files was preserved.	2022-05-27 19:33:35 +02:00
Willy Tarreau	a448227ba1	CLEANUP: quic: drop the name "conn_stream" from the pool variable names QUIC was the last user of entities with "conn_stream" in their names, though there's no more reason for this given that the pool names were already pretty straightforward. The renaming does this: qc_stream_desc: pool_head_quic_conn_stream -> pool_head_quic_stream_desc qc_stream_buf: pool_head_quic_conn_stream_buf -> pool_head_quic_stream_buf	2022-05-27 19:33:35 +02:00
Willy Tarreau	74568cf023	CLEANUP: stconn: rename final state manipulation functions from cs_* to sc_* This applies the following renaming. It's a bit large but pretty mechanical: cs_state -> sc_state (enum) cs_alloc_ibuf() -> sc_alloc_ibuf() cs_is_conn_error() -> sc_is_conn_error() cs_opposite() -> sc_opposite() cs_report_error() -> sc_report_error() cs_set_state() -> sc_set_state() cs_state_bit() -> sc_state_bit() cs_state_in() -> sc_state_in() cs_state_str() -> sc_state_str()	2022-05-27 19:33:35 +02:00
Willy Tarreau	f61dd19284	CLEANUP: stconn: rename cs_{shut,chk}* to sc_* This applies the following renaming: cs_shutr() -> sc_shutr() cs_shutw() -> sc_shutw() cs_chk_rcv() -> sc_chk_rcv() cs_chk_snd() -> sc_chk_snd() cs_must_kill_conn() -> sc_must_kill_conn()	2022-05-27 19:33:35 +02:00
Willy Tarreau	d68ff018c5	CLEANUP: stconn: rename cs{,_get}_{src,dst} to sc_* The following functions were renamed: cs_src() -> sc_src() cs_dst() -> sc_dst() cs_get_src() -> sc_get_src() cs_get_dst() -> sc_get_dst()	2022-05-27 19:33:35 +02:00
Willy Tarreau	19c65a9ded	CLEANUP: stconn: rename remaining management functions from cs_* to sc_* This is the end of the renaming for the generic SC management functions and macros: cs_applet_process() -> sc_applet_process() cs_attach_applet() -> sc_attach_applet() cs_attach_mux() -> sc_attach_mux() cs_attach_strm() -> sc_attach_strm() cs_detach_app() -> sc_detach_app() cs_detach_endp() -> sc_detach_endp() cs_notify() -> sc_notify() cs_reset_endp() -> sc_reset_endp() cs_state_in() -> sc_state_in() cs_update() -> sc_update() cs_update_rx() -> sc_update_rx() cs_update_tx() -> sc_update_tx() IS_HTX_CS() -> IS_HTX_SC()	2022-05-27 19:33:35 +02:00
Willy Tarreau	a0b58b537d	CLEANUP: stconn: rename cs_{new,create,free,destroy}_* to sc_* This renames the following functions: cs_new_from_endp() -> sc_new_from_endp() cs_new_from_strm() -> sc_new_from_strm() cs_new_from_check() -> sc_new_from_check() cs_applet_create() -> sc_applet_create() cs_destroy() -> sc_destroy() cs_free() -> sc_free()	2022-05-27 19:33:35 +02:00
Willy Tarreau	90e8b455b7	CLEANUP: stconn: rename cs_cant_get() to se_need_more_data() An equivalent applet_need_more_data() was added as well since that function is mostly used from applet code. It makes it much clearer that the applet is waiting for data from the stream layer.	2022-05-27 19:33:35 +02:00
Willy Tarreau	75a8f8e290	CLEANUP: stconn: rename cs_{want,stop}_get() to se_{will,wont}_consume() These ones are essentially for the stream endpoint, let's give them a name that matches the intent. Equivalent versions were provided in the applet namespace to ease code legibility.	2022-05-27 19:33:35 +02:00
Willy Tarreau	15252cd9c0	MEDIUM: stconn: move the RXBLK flags to the stream connector The following flags are not at all related to the endpoint but to the connector itself: - SE_FL_RXBLK_ROOM - SE_FL_RXBLK_BUFF - SE_FL_RXBLK_CHAN As such they have no business staying in the endpoint descriptor and they must move to the stream connector. They've also been renamed accordingly to better match what they correspond to (the same name as the function that sets them). The rare occurrences of cs_rx_blocked() were replaced by an explicit test on the list of flags. The reason is that cs_rx_blocked() used to preserve some tests that are not needed at certain places since already known. For the same reason SE_FL_RXBLK_ANY wasn't converted. As such it will later be possible to carefully review these few locations and eliminate the unneeded flags from the tests. No particular function was made to test them since they're explicit enough. It now looks like ci_putchk() and friends could very well place the flag themselves on the connector when they detect a buffer full condition, as this would significantly simplify the high-level API. But all usages must first be reviewed before this simplification can be done. For now it remains done by applet_put*() instead.	2022-05-27 19:33:35 +02:00
Willy Tarreau	8c02f8de14	CLEANUP: stconn: rename SE_FL_RX_WAIT_EP to SE_FL_HAVE_NO_DATA It's more explicit this way. The cs_rx_endp_ready() function could be removed so that the flag is directly tested. In the future it should be inverted and the few places where it's set (or preserved via SE_FL_APP_MASK) could be dropped.	2022-05-27 19:33:35 +02:00
Willy Tarreau	13d63afacd	MINOR: stconn: add sc_is_recv_allowed() to check for ability to receive At plenty of places we combine multiple flags checks to determine if we can receive (endp_ready, rx_blocked, cf_shutr etc). Let's group them under a single function that is meant to replace existing tests. Some tests were only checking the rxblk flags at the connection level, so for now they were not converted, this requires a bit of auditing first, and probably a test to determine whether or not to check for cf_shutr (e.g. there is none if no stream is present).	2022-05-27 19:33:35 +02:00
Willy Tarreau	4164eb94f3	MINOR: stconn: start to rename cs_rx_endp_{more,done}() to se_have_{no_,}more_data() The analysis of cs_rx_endp_more() showed that the purpose is for a stream endpoint to inform the connector that it's ready to deliver more data to that one, and conversely cs_rx_endp_done() that it's done delivering data so it should not be bothered again for this. This was modified two ways: - the operation is no longer performed on the connector but on the endpoint so that there is no more doubt when reading applet code about what this rx refers to; it's the endpoint that has more or no more data. - an applet implementation is also provided and mostly used from applet code since it saves the caller from having to access the endpoint descriptor. It's visible that the flag ought to be inverted because some places have to set it by default for no reason.	2022-05-27 19:33:35 +02:00
Willy Tarreau	0ed73c376c	CLEANUP: stconn: rename cs_rx_buff_{blk,rdy} to sc_{need,have}_buff() These functions are used by the application layer to disable or enable reading at the stream connector's level when the input buffer failed to be allocated (or was finally allocated). The new names makes things clearer.	2022-05-27 19:33:35 +02:00
Willy Tarreau	9512ab6e00	CLEANUP: stconn: rename cs_rx_chan_{blk,rdy} to sc_{wont,will}_read() These functions were used by the channel to inform the lower layer whether reading was acceptable or not. Usually this directly mimmicks the CF_DONT_READ flag from the channel, which may be set when it's desired not to buffer incoming data that will not be processed, or that the buffer wants to be flushed before starting to read again, or that bandwidth limiting might be enforced, etc. It's always a policy reason, not a purely resource-based one.	2022-05-27 19:33:35 +02:00
Willy Tarreau	99615ed85d	CLEANUP: stconn: rename cs_rx_room_{blk,rdy} to sc_{need,have}_room() The new name mor eclearly indicates that a stream connector cannot make any more progress because it needs room in the channel buffer, or that it may be unblocked because the buffer now has more room available. The testing function is sc_waiting_room(). This is mostly used by applets. Note that the flags will change soon.	2022-05-27 19:33:35 +02:00
Willy Tarreau	b73262fc85	MEDIUM: stconn: take SE_FL_APPLET_NEED_CONN out of the RXBLK_ANY flags This makes SE_FL_APPLET_NEED_CONN autonomous, in that we check for it everywhere we have a relevant cs_rx_blocked(), so that the flag doesn't need anymore to be covered by cs_rx_blocked(). Indeed, this flag doesn't really translate a receive blocking condition but rather a refusal to wake up an applet that is waiting for a connection to finish to setup. This also ensures we will not risk to set it back on a new endpoint after cs_reset_endp() via SE_FL_APP_MASK, because the flag being specific to the endpoint only and not to the connector, we don't want to preserve it when replacing the endpoint. It's possible that cs_chk_rcv() could later be further simplified if we can demonstrate that the two tests in it can be merged.	2022-05-27 19:33:35 +02:00
Willy Tarreau	b23edc8b8d	MINOR: stconn: rename SE_FL_RXBLK_CONN to SE_FL_APPLET_NEED_CONN This flag is exclusively used when a front applet needs to wait for the other side to connect (or fail to). Let's give it a more explicit name and remove the ambiguous function that was used only once. This also ensures we will not risk to set it back on a new endpoint after cs_reset_endp() via SE_FL_APP_MASK, because the flag being specific to the endpoint only and not to the connector, we don't want to preserve it when replacing the endpoint.	2022-05-27 19:33:35 +02:00
Willy Tarreau	676c8db134	MEDIUM: stconn: remove SE_FL_RXBLK_SHUT This flag is no more needed, it was only set on shut read to be tested by cs_rx_blocked() which is now properly tested for shutr as well. The cs_rx_blk_shut() calls were removed. Interestingly it allowed to remove a special case in the L7 retry code. This also ensures we will not risk to set it back on a new endpoint after cs_reset_endp() via SE_FL_APP_MASK.	2022-05-27 19:33:35 +02:00
Willy Tarreau	e7866b1ff7	MEDIUM: stconn: always rely on CF_SHUTR in addition to cs_rx_blocked() One flag (RXBLK_SHUT) is always set with CF_SHUTR, so in order to remove it, we first need to make sure we always check for CF_SHUTR where cs_rx_blocked() is being used.	2022-05-27 19:33:35 +02:00
Willy Tarreau	516621bbe6	MINOR: stconn: remove calls to cs_done_get() It was only called after setting SHUTW on the output channel, and since it's now handled by sc_is_send_allowed() we don't need it anymore.	2022-05-27 19:33:34 +02:00
Willy Tarreau	902ba7e2bc	CLEANUP: stconn: use a single function to know if SC may send to SE sc_is_send_allowed() is now used everywhere instead of the combination of cs_tx_endp_ready() && !cs_tx_blocked(). There's no place where we need them individually thus it's simpler. The test was placed in cs_util as we'll complete it later.	2022-05-27 19:33:34 +02:00
Willy Tarreau	967955b156	CLEANUP: stconn: rename cs_ep_set_error() to se_fl_set_error() First it applies to the stream endpoint and not the conn_stream, and second it only tests and touches the flags so it makes sense to call it se_fl_ like other functions which only manipulate the flags, as it's just a special case of flags.	2022-05-27 19:33:34 +02:00
Willy Tarreau	108423819c	CLEANUP: stconn: rename cs_conn_get_first() to conn_get_first_sc() It returns an stconn from a connection and not the opposite, so the name change was more appropriate. In addition it was moved to connection.h which manipulates the connection stuff, and it happens that only connection.c uses it.	2022-05-27 19:33:34 +02:00
Willy Tarreau	462b989d4c	CLEANUP: stconn: rename cs_conn_() to sc_conn_() The following functions which act on a connection-based stream connector were renamed to sc_conn_* (~60 places): cs_conn_drain_and_shut cs_conn_process cs_conn_read0 cs_conn_ready cs_conn_recv cs_conn_send cs_conn_shut cs_conn_shutr cs_conn_shutw	2022-05-27 19:33:34 +02:00
Willy Tarreau	f8d0ab54ec	CLEANUP: stconn: rename cs_get_data_name() to sc_get_data_name() Only used twice to dump stream debug info.	2022-05-27 19:33:34 +02:00
Willy Tarreau	fa57cc7b20	CLEANUP: stconn: rename __cs_endp_target() to __sc_endp() The function returns the real stream endpoint so since there's no more confusion around the terminology, let's drop "target".	2022-05-27 19:33:34 +02:00
Willy Tarreau	8e7c6e6907	CLEANUP: stconn: rename cs_appctx() to sc_appctx() Nothing special, just s/cs/sc/, roughly 50-60 entries.	2022-05-27 19:33:34 +02:00
Willy Tarreau	417a31bb55	CLEANUP: stconn: rename cs_conn_mux() to sc_mux_ops() This effectively returns the mux_ops from the connection when it exists on an stconn.	2022-05-27 19:33:34 +02:00
Willy Tarreau	6fe2b42e45	CLEANUP: stconn: rename cs_mux() to sc_mux_strm() The function doesn't return a pointer to the mux but to the mux stream (h1s, h2s etc). Let's adjust its name to reflect this. It's rarely used, the name can be enlarged a bit. And of course s/cs/sc to accommodate for the updated name.	2022-05-27 19:33:34 +02:00
Willy Tarreau	fd9417ba3f	CLEANUP: stconn: rename cs_conn() to sc_conn() It's mostly used from upper layers. Both the checked and unchecked functions were updated, or ~150 entries.	2022-05-27 19:33:34 +02:00
Willy Tarreau	ea27f48c5a	CLEANUP: stconn: rename cs_{check,strm,strm_task} to sc_strm_* These functions return the app-layer associated with an stconn, which is a check, a stream or a stream's task. They're used a lot to access channels, flags and for waking up tasks. Let's just name them appropriately for the stream connector.	2022-05-27 19:33:34 +02:00
Willy Tarreau	40a9c32e3a	CLEANUP: stconn: rename cs_{i,o}{b,c} to sc_{i,o}{b,c} We're starting to propagate the stream connector's new name through the API. Most call places of these functions that retrieve the channel or its buffer are in applets. The local variable names are not changed in order to keep the changes small and reviewable. There were ~92 uses of cs_ic(), ~96 of cs_oc() (due to co_get() being less factorizable than ci_put), and ~5 accesses to the buffer itself.	2022-05-27 19:33:34 +02:00
Willy Tarreau	d0a06d52f4	CLEANUP: applet: use applet_put() everywhere possible This applies the change so that the applet code stops using ci_putchk() and friends everywhere possible, for the much saferapplet_put() instead. The change is mechanical but large. Two or three functions used to have no appctx and a cs derived from the appctx instead, which was a reminiscence of old times' stream_interface. These were simply changed to directly take the appctx. No sensitive change was performed, and the old (more complex) API is still usable when needed (e.g. the channel is already known). The change touched roughly a hundred of locations, with no less than 124 lines removed. It's worth noting that the stats applet, the oldest of the series, could get a serious lifting, as it's still very channel-centric instead of propagating the appctx along the chain. Given that this code doesn't change often, there's no emergency to clean it up but it would look better.	2022-05-27 19:33:34 +02:00
Willy Tarreau	2f2318df87	MEDIUM: stconn: merge the app_ops and the data_cb fields For historical reasons (stream-interface and connections), we used to require two independent fields for the application level callbacks and the transport-level functions. Over time the distinction faded away so much that the low-level functions became specific to the application and conversely. For example, applets may only work with streams on top since they rely on the channels, and the stream-level functions differ between applets and connections. Right now the application level only contains a wake() callback and the low-level ones contain the functions that act at the lower level to perform the shutr/shutw and at the upper level to notify about readability and writability. Let's just merge them together into a single set and get rid of this confusing distinction. Note that the check ops do not define any app-level function since these are only called by streams.	2022-05-27 19:33:34 +02:00
Willy Tarreau	f3ae34b67d	MINOR: check: export wake_srv_chk() We'll need it to centralize the stream connectors definitions.	2022-05-27 19:33:34 +02:00
Willy Tarreau	026e8fb290	CLEANUP: stconn: tree-wide rename stconn states CS_ST/SB_* to SC_ST/SB_* This also follows the natural naming. There are roughly 238 changes, all totally trivial. conn_stream-t.h has become completely void of any "conn_stream" related stuff now (except its name).	2022-05-27 19:33:34 +02:00
Willy Tarreau	cb04166525	CLEANUP: stconn: tree-wide rename stream connector flags CS_FL_* to SC_FL_* This follows the natural naming. There are roughly 100 changes, all totally trivial.	2022-05-27 19:33:34 +02:00
Willy Tarreau	7cb9e6c6ba	CLEANUP: stream: rename "csf" and "csb" to "scf" and "scb" These are the stream connectors, let's give them consistent names. The patch is large (405 locations) but totally trivial.	2022-05-27 19:33:34 +02:00
Willy Tarreau	c105492bf5	CLEANUP: stdesc: rename the stream connector ->cs field to ->sc This is a rename of this field. Most of the places were in muxes, but were already factored with the previous series adding *_sc().	2022-05-27 19:33:34 +02:00
Willy Tarreau	32c095b622	CLEANUP: mux-pt: add and use pt_sc() to retrieve the stream connector This is better and easier to adapt than pt->endp->cs.	2022-05-27 19:33:34 +02:00
Willy Tarreau	77534276d1	CLEANUP: mux-fcgi: add and use fcgi_strm_sc() to retrieve the stream connector This is better and easier to adapt than fstrm->endp->cs.	2022-05-27 19:33:34 +02:00
Willy Tarreau	7be4ee0673	CLEANUP: mux-h2: add and use h2s_sc() to retrieve the stream connector This is better and easier to adapt than h2s->endp->cs.	2022-05-27 19:33:34 +02:00
Willy Tarreau	97b4d3bc36	CLEANUP: mux-h1: add and use h1s_sc() to retrieve the stream connector This is better and easier to adapt than h1s->endp->cs.	2022-05-27 19:33:34 +02:00
Willy Tarreau	4596fe20d9	CLEANUP: conn_stream: tree-wide rename to stconn (stream connector) This renames the "struct conn_stream" to "struct stconn" and updates the descriptions in all comments (and the rare help descriptions) to "stream connector" or "connector". This touches a lot of files but the change is minimal. The local variables were not even renamed, so there's still a lot of "cs" everywhere.	2022-05-27 19:33:34 +02:00
Willy Tarreau	3a3f480d15	CLEANUP: conn_stream: rename cs_app_* to sc_app_* Let's start to introduce the stream connector at the app_ops level. This is entirely self-contained into conn_stream.c. The functions were also updated to reflect the new name, and the comments were updated.	2022-05-27 19:33:34 +02:00
Willy Tarreau	798465b02c	CLEANUP: conn_stream: rename the conn_stream's endp to sedesc Just like for the appctx, this is a pointer to a stream endpoint descriptor, so let's make this explicit and not confuse it with the full endpoint. There are very few changes thanks to the preliminary refactoring of the flags manipulation.	2022-05-27 19:33:34 +02:00
Willy Tarreau	d869e13ed8	CLEANUP: applet: rename the sedesc pointer from "endp" to "sedesc" Now at least it makes it obvious that it's the stream endpoint descriptor and not an endpoint. There were few changes thanks to the previous refactor of the flags.	2022-05-27 19:33:34 +02:00
Willy Tarreau	ea59b0201c	CLEANUP: conn_stream: rename cs_endpoint to sedesc (stream endpoint descriptor) After some discussion we found that the cs_endpoint was precisely the descriptor for a stream endpoint, hence the naturally coming name, stream endpoint constructor. This patch renames only the type everywhere and the new/init/free functions to remain consistent with it. Future patches will address field names and argument names in various code areas.	2022-05-27 19:33:34 +02:00
Willy Tarreau	65d0597b2b	CLEANUP: conn_stream: rename the cs_endpoint's target to "se" That's the "stream endpoint" pointer. Let's change it now while it's not much spread. The function __cs_endp_target() wasn't yet renamed because that will change more globally soon.	2022-05-27 19:33:34 +02:00
Willy Tarreau	b605c4213f	CLEANUP: conn_stream: rename the stream endpoint flags CS_EP_* to SE_FL_* Let's now use the new flag names for the stream endpoint.	2022-05-27 19:33:34 +02:00
Willy Tarreau	d56377c5eb	CLEANUP: conn_stream: apply endp_flags.cocci tree-wide This changes all main uses of endp->flags to the se_fl_() equivalent by applying coccinelle script endp_flags.cocci. The se_fl_() functions themselves were manually excluded from the change, of course. Note: 144 locations were touched, manually reviewed and found to be OK. The script was applied with all includes: spatch --in-place --recursive-includes -I include --sp-file $script $files	2022-05-27 19:33:34 +02:00
Willy Tarreau	0cfcc40812	CLEANUP: conn_stream: apply cs_endp_flags.cocci tree-wide This changes all main uses of cs->endp->flags to the sc_ep_*() equivalent by applying coccinelle script cs_endp_flags.cocci. Note: 143 locations were touched, manually reviewed and found to be OK, except a single one that was adjusted in cs_reset_endp() where the flags are read and filtered to be used as-is and not as a boolean, hence was replaced with sc_ep_get() & $FLAGS. The script was applied with all includes: spatch --in-place --recursive-includes -I include --sp-file $script $files	2022-05-27 19:33:34 +02:00
Willy Tarreau	24d15b1891	CLEANUP: conn_stream: rename the cs_endpoint's context to "conn" This one is exclusively used by the connection, regardless its generic name "ctx" is rather confusing. Let's make it a struct connection* and call it "conn". This way there's no doubt about what it is and there's no way it will be used by accident by being taken for something else.	2022-05-27 19:33:34 +02:00
Willy Tarreau	5fec7a1f98	CLEANUP: conn_stream: remove unneeded exclusion of RX_WAIT_EP from RXBLK_ANY This test in cs_update_rx() was introduced in 1.9 by commit `b26a6f970` ("MEDIUM: stream-int: make use of si_rx_chan_{rdy,blk} to control the stream-int from the channel"), but by then already it was not needed because the RX_WAIT_EP flag has never been part of RXBLK_ANY so there's no point doing "flags & RXBLK_ANY & ~RX_WAIT_EP", that part is already complicated enough like this.	2022-05-27 19:33:34 +02:00
Thayne McCombs	6a0d217628	BUG/MEDIUM: sample: Fix adjusting size in word converter Adjust the size of the sample buffer before we change the "area" pointer. Otherwise, we end up not changing the size, because the area pointer is already the same as "start" before we compute the difference between the two. This is similar to the change in `b28430591d` but for the word converter instead of field.	2022-05-27 19:33:34 +02:00
William Lallemand	d8c195a326	BUG/MINOR: ssl/lua: use correctly cert_ext in CertCache.set() Fix a typo that lead to using the wrong pointer when loading a certificate, which lead to always using the pem loader for every parameeter. Use the cert_ext->load() ptr instead of cert_exts->load() which was the first element of the cert_exts[] array. Enhance the error message with the field name. Should fix issue #1716	2022-05-26 19:36:07 +02:00
Willy Tarreau	8e5b9589b3	CLEANUP: init: address another coverity warning about a possible multiply overflow Commit `2cb3be76b` ("CLEANUP: init: address a coverity warning about possible multiply overflow") was incomplete, two other locations were present. This should address issue #1585.	2022-05-26 08:55:05 +02:00
Amaury Denoyelle	8c6176b8db	MINOR: h3: refactor SETTINGS parsing/error reporting Bring some improvment to h3_parse_settings_frm() function. The first one is the parsing which now manipulates a buffer instead of a plain char. This is more to unify with other parsing functions rather than dealing with data wrapping : it's unlikely to happen as SETTINGS is only received as the first frame on the control STREAM. Various errors are now properly reported as connection error : on incomplete frame payload * on a duplicated settings in the same frame * on reserved settings receive	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	849b24f15b	MINOR: h3: abort read on unknown uni stream As specified by HTTP/3 draft, an unknown unidirectional stream can be aborted. To do this, use a new flag QC_SF_READ_ABORTED. When the MUX detects this flag, QCS instance is automatically freed. Previously, such streams were instead automatically drained. By aborting them, we economize some useless memcpy instruction. On future data reception, QCS instance is not found in the tree and considered as already closed. The frame payload is thus deleted without copying it.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	9cc475182c	CLEANUP: h3: remove h3 uni tasklet Remove all unnecessary bits of code for H3 unidirectional streams. Most notable, an individual tasklet is not require anymore for each stream. This is useless since the merge of RX/TX uni streams handling with bidirectional streams code.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	f8db5aaf78	MEDIUM: quic: refactor uni streams RX The whole QUIC stack is impacted by this change : * at quic-conn level, a single function is now used to handle uni and bidirectional streams. It uses qcc_recv() function from MUX. * at MUX level, qc_recv() io-handler function does not skip uni streams * most changes are conducted at app layer. Most notably, all received data is handle by decode_qcs operation. Now that decode_qcs is the single app read function, the H3 layer can be simplified. Uni streams parsing was extracted from h3_attach_ruqs() to h3_decode_qcs(). h3_decode_qcs() is able to deal with all HTTP/3 frame types. It first check if the frame is valid for the H3 stream type. Most notably, SETTINGS parsing was moved from h3_control_recv() into h3_decode_qcs(). This commit has some major benefits besides removing duplicated code. Mainly, QUIC flow control is now enforced for uni streams as with bidi streams. Also, an unknown frame received on control stream does not set an error : it is now silently ignored as required by the specification. Some cleaning in H3 code is already done with this patch : h3_control_recv() and h3_attach_ruqs() are removed as they are now unused. A final patch should clean up the unneeded remaining bit.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	fc99a6933e	MINOR: h3: define non-h3 generic parsing function Define a new function h3_parse_uni_stream_no_h3(). It can be used to handle the payload of streams which does not convey H3 frames. This is mainly useful for QPACK encoder/decoder streams. It can also be used for a stream of unknown type which should be drain without parsing it. This patch is useful to extract code in a dedicated function. It will be simple to reuse it in h3_decode_qcs() when uni-streams reception is unify with bidirectional streams, without using dedicated stream tasklet.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	302ecd490b	MINOR: h3: check if frame is valid for stream type Define a new function h3_is_frame_valid(). It returns if a frame is valid or not depending on the stream which received it. For the moment, it is used in h3_decode_qcs() which only deals with bidirectional streams. Soon, uni streams will use the same function, rendering the frame type check useful.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	3555064e87	MINOR: h3: refactor uni streams initialization Define a new function h3_init_uni_stream(). This can be used to read the stream type of an unidirectional stream. There is no functional change with previous code. This patch will be useful to unify reception for uni streams with bidirectional ones.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	3236a8e85c	MINOR: h3: define stream type Define a new enum h3s_t. This is used to differentiate between the different stream types used in a HTTP/3 connection, including the QPACK encoder/decoder streams. For the moment, only bidirectional streams is positioned. This patch will be useful to unify reception of uni streams with bidirectional ones.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	6b92394973	MINOR: h3/qpack: use qcs as type in decode callbacks Replace h3_uqs type by qcs in stream callbacks. This change is done in the context of unification between bidi and uni-streams. h3_uqs type will be unneeded when this is achieved.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	c6195d77b4	BUG/MINOR: mux-quic: refactor uni streams TX/send H3 SETTINGS Remove the unneeded skip over unidirectional streams in qc_send(). This unify sending for both uni and bidi streams. In fact, the only local unidirectional streams in use for the moment is the H3 Control stream responsible of SETTINGS emission. The frame was already properly generated in qcs.tx.buf, but not send due to stream skip in qc_send(). Now, there is no need to ignore uni streams so remove this condition. This fixes the emission of H3 settings which is now properly emitted. Uni and bidi streams use the same set of funtcions for sending. One of the most notable gain is that flow-control is now enforced for uni streams.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	6754d7e2ed	MINOR: mux-quic: emit STREAM_STATE_ERROR in qcc_recv Emit STREAM_STATE_ERROR connection error in two cases : * if receiving data for send-only stream * if receiving data on a locally initiated stream not open yet For the moment the first case cannot be encoutered as uni streams reception does not use qcc_recv(). However, this will be soon implemented with the unification between bidi and uni streams.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	80097cc824	MINOR: h3: reject too big frames The whole frame payload must have been received to demux a H3 frames, except for H3 DATA which can be fragmented into multiple HTX blocks. If the frame is bigger than the buffer and is not a DATA frame, a connection error is reported with error H3_EXCESSIVE_LOAD. This should be completed in the future with the H3 settings to limit the size of uncompressed header section. This code is more generic : it can handle every H3 frames. This is done in order to be able to use h3_decode_qcs() to demux both uni and bidir streams.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	5c4373a47b	MINOR: mux-quic: disable read on CONNECTION_CLOSE emission Similar to sending, read operations are disabled when a CONNECTION_CLOSE frame has been emitted. Most notably, this prevents unneeded loop demuxing when the H3 layer has issue an error and cannot process the buffer payload anymore. Note that read is not prevented for unidirectional streams for the moment. This will supported soon with the unification of bidir and uni streams treatment.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	f9e190e49a	MINOR: quic: support CONNECTION_CLOSE_APP emission Complete quic-conn API for error reporting. A new parameter <app> is defined in the function quic_set_connection_close(). This will transform the frame into a CONNECTION_CLOSE_APP type. This type of frame will be generated by the applicative layer, h3 or hq-interop for the moment. A new function qcc_emit_cc_app() is exported by the MUX layer for them.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	65df3add33	MINOR: h3: refactor h3_control_send() The only change is that the H3_CF_SETTINGS_SENT flag if-condition is replaced by a BUG_ON statement. This may help to catch multiple calls on h3_control_send() instead of silently ignore them.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	160507d0ba	BUG/MINOR: h3: prevent overflow when parsing SETTINGS h3_parse_settings_frm() read one byte after the frame payload. Fix the parsing code. In most cases, this has no impact as we are inside an allocated buffer but it could cause a segfault depending on the buffer alignment.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	081479df92	CLEANUP: h3: rename uni stream type constants Cosmetic fix which reduce the name of unidirectional stream constants. No impact on the code.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	8d1ecac5d9	CLEANUP: h3: rename struct h3 -> h3c struct h3 represents the whole HTTP/3 connection. A new type h3s was recently introduced to represent a single HTTP/3 stream. To facilitate the analogy with other haproxy code, most notable in MUX, rename h3 type to h3c.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	0ffd6e7e64	MINOR: mux-quic: adjust return value of decode_qcs Use 0 for success of decode_qcs operation else non-zero. This is to follow the same model which is in use in most of the function in MUX/H3 code.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	e1cad8bc03	MINOR: mux-quic: add traces in qc_recv() Just add traces in qc_recv() similarly to qc_send() function.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	1c25b18e17	MINOR: mux-quic: delay cs_endpoint allocation Do not allocate cs_endpoint for every QCS instances in qcs_new(). Instead, this is delayed to qc_attach_cs() function. In effect, with H3 as app protocol, cs_endpoint will be allocated on HEADERS parsing. Thus, no cs_endpoint is allocated for H3 unidirectional streams which do not convey any HTTP data.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	93fba32430	MINOR: mux-quic: do not alloc quic_stream_desc for uni remote stream qc_stream_desc type is required for sending. Thus, it is not required for an unidirectional remote stream where only receive will be performed.	2022-05-25 15:41:25 +02:00
Amaury Denoyelle	c7dd9d6867	MINOR: h3: mark ncbuf as const on h3_b_dup h3_b_dup() is used to obtains a ncbuf representation into a struct buffer. ncbuf can thus be marked as a const parameter. This will allows function which already manipulates a const ncbuf to use it.	2022-05-25 15:41:25 +02:00
Emeric Brun	ca82578fe8	BUG/MEDIUM: peers: prevent unitialized multiple listeners on peers section The previous fix: BUG/MEDIUM: peers: fix segfault using multiple bind on peers Prevents to declare multiple listeners on a peers sections but if peers protocol is extended to support this we could raise the bug again. Indeed, after allocating a new listener and adding it to a list the code mistakenly re-configure the first element of the list instead of the new added one, and the last one remains finally uninitialized. The previous fix assure there is no more than one listener in this list but this could be changed in futur. This patch patch assures we configure and initialize the newly added listener instead of the first one in the list. This patch could be backported until version 2.0 to complete BUG/MEDIUM: peers: fix segfault using multiple bind on peers	2022-05-25 15:10:08 +02:00
Emeric Brun	49f6f4b1a7	BUG/MEDIUM: peers: fix segfault using multiple bind on peers sections If multiple "bind" lines were present on the "peers" section, multiple listeners were added to a list but the code mistakenly initialize the first member and this first listener was re-configured instead of the newly created one. The last one remains uninitialized causing a null dereference a soon a connection is received. In addition, the 'peers' sections and protocol are not currently designed to handle multiple listeners. This patch check if there is already a listener configured on the 'peers' section when we want to create a new one. This is rising an error if a listener is already present showing the file and line in the error message. To keep the file and line number of the previous listener available for the error message, the 'bind_conf_uniq_alloc' function was modified to keep the file/line data the struct 'bind_conf' was firstly allocated (previously it was updated each time the 'bind_conf' was reused). This patch should be backported until version 2.0	2022-05-25 15:10:08 +02:00
Christopher Faulet	4315d17d3f	BUG/MEDIUM: resolvers: Don't defer resolutions release in deinit function resolvers_deinit() function is called on error, during post-parsing stage, or on deinit, when HAProxy is stopped. It releases all entities: resolvers, resolutions and SRV requests. There is no reason to defer the resolutions release by moving them in the death_row list because this function is terminal. And it is in fact a bug. Resolutions must not be released at the end of the function because resolvers were already freed. However some resolutions may still be attached to a reolver. Thus, when we try to remove it from the resolver's tree, in resolv_reset_resolution(), this resolver was already released. So now, resolution are immediately released. It means there is no more reason to track this function. calls to enter_resolver_code()/leave_resolver_code() have been removed. This patch should fix the issue #1680 and may be related to #1485. It must be backported as far as 2.2.	2022-05-24 18:11:59 +02:00
Willy Tarreau	1ba30167a0	MEDIUM: h1: enlarge the scope of accepted version chars with accept-invalid-http-request We used to support both RTSP and HTTP protocol version names with and without accept-invalid-http-request, but since this is based on the characters themselves, any protocol made of chars {0-9/.HPRST} was possible and not others. Now that such non-standard protocols are restricted to accept-invalid-http-request, there's no reason for not allowing other letters. With this patch, characters {0-9./A-Z} are permitted when the option is set.	2022-05-24 15:38:54 +02:00
Tim Duesterhus	8f4116ea65	BUG/MEDIUM: http: Properly reject non-HTTP/1.x protocols This patch hardens the verification of the HTTP/1.x version line (i.e. the first line within an HTTP/1.x request) to verify that the protocol name within the version actually reads "HTTP". Previously protocols that superficially resembled the wire-format of HTTP/1.x and having a 4-letter acronym as the protocol name, such as RTSP would pass this check. This patch fixes GitHub issue #540, it must be backported to all supported versions. The legacy, non-HTX parser is affected as well, a fix must be created for it as well. Note that such protocols can still be used when option accept-invalid-http-request is set.	2022-05-24 15:38:05 +02:00
Willy Tarreau	2cb3be76bf	CLEANUP: init: address a coverity warning about possible multiply overflow In issue #1585 Coverity suspects a risk of multiply overflow when calculating the SSL cache size, though in practice the cache is limited to 2^32 anyway thus it cannot really happen. Nevertheless, casting the operation should be sufficient to avoid marking it as a false positive.	2022-05-24 07:46:00 +02:00
Amaury Denoyelle	019a182ebf	Revert "MINOR: mux-quic: activate qmux traces on stdout via macro" This reverts commit `251eadfce5`. This patch is similar to the previous revert for QUIC-MUX traces.	2022-05-23 11:06:42 +02:00
Amaury Denoyelle	2208d57ce7	Revert "MINOR: quic: activate QUIC traces at compilation" This reverts commit `118b2cbf84`. This patch was useful mainly for the docker image of QUIC interop to have traces on stdout. A better solution has been found by integrating this patch directly in the qns repository which is used to build the docker image. Thus, this hack is not require anymore in the main repository.	2022-05-23 11:06:42 +02:00
Amaury Denoyelle	d0c62328af	BUG/MEDIUM: mux-quic: adjust buggy proxy closing support The wake handler detects if the frontend is closed. This can happen if the proxy has been disabled individually or even on process soft-stop. Before this patch, in this condition QCS instances were freed before being detached from the cs_endpoint. This clearly violates the haproxy connection architecture and cause a BUG_ON statement crash in cs_free(). To handle this properly, cs_endpoint is notified by setting RD_SH\|WR_SH on connection flags. The cs_endpoint will thus use the detach operation which allows the QCS instance to be freed. This code allows the soft-stop process to complete as soon as possible. However, the client is not notified about the connection closing. It should be done by emitting a H3 GOAWAY + CONNECTION_CLOSE. Sadly, this is impossible at this stage because the listener sockets are closed so the quic-conn cannot use it to emit new frames. At this stage the client will most probably detect connection closing on its idle timeout expiration. Thus, to completely support proxy closing/soft-stop, important architecture changes are required in QUIC socket management. This is also linked with the reload feature.	2022-05-23 11:05:52 +02:00
Tim Duesterhus	22535a56a7	CLEANUP: tools: Crash if inet_ntop fails due to ENOSPC in sa2str This is impossible, because we pass a destination buffer that is appropriately sized to hold an IPv6 address. This is related to GitHub issue #1599.	2022-05-23 09:39:32 +02:00
Tim Duesterhus	162f0875ad	BUG/MEDIUM: tools: Fix `inet_ntop` usage in sa2str The given size must be the size of the destination buffer, not the size of the (binary) address representation. This fixes GitHub issue #1599. The bug was introduced in `92149f9a82` which is in 2.4+. The fix must be backported there.	2022-05-23 08:45:31 +02:00
Tim Duesterhus	147eeb2ef3	CLEANUP: tools: Clean up non-QUIC error message handling in str2sa_range() If QUIC support is enabled both branches of the ternary conditional are identical, upsetting Coverity. Move the full conditional into the non-QUIC preprocessor branch to make the code more clear. This resolves GitHub issue #1710.	2022-05-23 08:45:31 +02:00
Frédéric Lécaille	07968dcea3	BUG/MINOR: quic: Missing <conn_opening> stats counter decrementation When we receive a CONNECTION_CLOSE frame, we should decrement this counter if the handshake state was not successful and if we have not received a TLS alert from the TLS stack.	2022-05-20 20:57:31 +02:00
Frédéric Lécaille	ad9895d133	BUG/MINOR: quic: Fixe a typo in qc_idle_timer_task() The & operator was confused with \| operator :-(	2022-05-20 20:57:31 +02:00
Willy Tarreau	287f32fd01	MINOR: listener: automatically enable SSL if a QUIC transport is found When a bind line is configured without the "ssl" keyword, a warning is emitted and a crash happens at runtime: bind quic4@:4449 crt rsa+dh2048.pem alpn h3 allow-0rtt [WARNING] (17867) : config : Proxy 'decrypt': A certificate was specified but SSL was not enabled on bind 'quic4@:4449' at [quic-mini.cfg:24] (use 'ssl'). Let's automatically turn SSL on when QUIC is detected, as it doesn't exist without SSL anyway. It solves the runtime issue, and also makes sure it is not possible to accidentally configure a quic listener with no certificate since the error is detected via the SSL checks. A warning is emitted in this case, to encourage the user to fix the configuration so that it remains reviewable.	2022-05-20 18:41:55 +02:00
Willy Tarreau	730cc02c26	MINOR: listener: automatically select a QUIC mux with a QUIC transport When no mux protocol is configured on a bind line with "proto", and the transport layer is QUIC, right now mux_h1 is being used, leading to a crash. Now when the transport layer of the bind line is already known as being QUIC, let's automatically try to configure the QUIC mux, so that users do not have to enter "proto quic" all the time while it's the only supported option. this means that the following line now works: bind quic4@:4449 ssl crt rsa+dh2048.pem alpn h3 allow-0rtt	2022-05-20 18:41:55 +02:00
Willy Tarreau	cac619b4d2	MINOR: config: detect and report mux and transport incompatibilities Till now, placing "proto h1" or "proto h2" on a "quic" bind or placing "proto quic" on a TCP line would parse fine but would crash when traffic arrived. The reason is that there's a strong binding between the QUIC mux and QUIC transport and that they're not expected to be called with other types at all. Now that we have the mux's type and we know the type of the protocol used on the bind conf, we can perform such checks. This now returns: [ALERT] (16978) : config : frontend 'decrypt' : stream-based MUX protocol 'h2' is incompatible with framed transport of 'bind quic4@:4448' at [quic-mini.cfg:27]. [ALERT] (16978) : config : frontend 'decrypt' : frame-based MUX protocol 'quic' is incompatible with stream transport of 'bind :4448' at [quic-mini.cfg:29]. This config tightening is only tagged MINOR since while such a config, despite not reporting error, cannot work at all so even if it breaks experimental configs, they were just waiting for a single connection to crash.	2022-05-20 18:41:55 +02:00
Willy Tarreau	b5821e12ce	MINOR: connection: add flag MX_FL_FRAMED to mark muxes relying on framed xprt In order to be able to check compatibility between muxes and transport layers, we'll need a new flag to tag muxes that work on framed transport layers like QUIC. Only QUIC has this flag now.	2022-05-20 18:41:55 +02:00
Willy Tarreau	2071a99dfe	MINOR: listener/ssl: set the SSL xprt layer only once the whole config is known We used to preset XPRT_SSL on bind_conf->xprt when parsing the "ssl" keyword, which required to be careful about what QUIC could have set before, and which makes it impossible to consider the whole line to set all options. Now that we have the BC_O_USE_SSL option on the bind_conf, it becomes easier to set XPRT_SSL only once the bind_conf's args are parsed.	2022-05-20 18:41:55 +02:00
Willy Tarreau	78d0dcd519	MINOR: listener: set the QUIC xprt layer immediately after parsing the args It used to be set when parsing the listeners' addresses but this comes with some difficulties in that other places have to be careful not to replace it (e.g. the "ssl" keyword parser). Now we know what protocols a bind_conf line relies on, we can set it after having parsed the whole line.	2022-05-20 18:41:55 +02:00
Willy Tarreau	64306ccd97	MINOR: listener: detect stream vs dgram conflict during parsing Now that we have a function to parse all bind keywords, and that we know what types of sock-level and xprt-level protocols a bind_conf is using, it's easier to centralize the check for stream vs dgram conflict by putting it directly at the end of the args parser. This way it also works for peers, provides better precision in the report, and will also allow to validate transport layers. The check was even extended to detect inconsistencies between xprt layer (which were not covered before). It can even detect that there are two incompatible "bind" lines in a single peers section.	2022-05-20 18:41:55 +02:00
Willy Tarreau	91b780a455	CLEANUP: listener: store stream vs dgram at the bind_conf level Let's collect the set of xprt-level and sock-level dgram/stream protocols seen on a bind line and store that in the bind_conf itself while they're being parsed. This will make it much easier to detect incompatibilities later than the current approch which consists in scanning all listeners in post-parsing.	2022-05-20 18:41:55 +02:00
Willy Tarreau	787e92a4fb	CLEANUP: listener: replace bind_conf->quic_force_retry with BC_O_QUIC_FORCE_RETRY It was only set and used once, let's replace it now and take it out of the ifdef.	2022-05-20 18:41:51 +02:00
Willy Tarreau	1ea6e6a17f	CLEANUP: listener: replace bind_conf->generate_cers with BC_O_GENERATE_CERTS The new flag will now replace this boolean variable.	2022-05-20 18:39:43 +02:00
Willy Tarreau	11ba404c6b	CLEANUP: listener: replace all uses of bind_conf->is_ssl with BC_O_USE_SSL The new flag will now replace this boolean variable that was only set and tested.	2022-05-20 18:39:43 +02:00
Willy Tarreau	55f0f7bb54	MINOR: config: use the new bind_parse_args_list() to parse a "bind" line This now makes sure that both the peers' "bind" line and the regular one will use the exact same parser with the exact same behavior. Note that the parser applies after the address and that it could be factored further, since the peers one still does quite a bit of duplicated work.	2022-05-20 18:39:43 +02:00
Willy Tarreau	3882d2a96c	MINOR: listener: provide a function to process all of a bind_conf's arguments The "bind" parsing code was duplicated for the peers section and as a result it wasn't kept updated, resulting in slightly different error behavior (e.g. errors were not freed, warnings were emitted as alerts) Let's first unify it into a new dedicated function that properly reports and frees the error.	2022-05-20 18:39:43 +02:00
Willy Tarreau	91b47263f7	MINOR: protocol: replace ctrl_type with xprt_type and clarify it There's been some great confusion between proto_type, ctrl_type and sock_type. It turns out that ctrl_type was improperly chosen because it's not the control layer that is of this or that type, but the transport layer, and it turns out that the transport layer doesn't (normally) denaturate the underlying control layer, except for QUIC which turns dgrams to streams. The fact that the SOCK_{DGRAM\|STREAM} set of values was used added to the confusion. Let's replace it with xprt_type which reuses the later introduced PROTO_TYPE_* values, and update the comments to explain which one works at what level.	2022-05-20 18:39:43 +02:00
Willy Tarreau	3d7b4684fe	CLEANUP: config: provide cleare hints about unsupported QUIC addresses We now detect that QUIC was likely requested, and if it's not compiled it, we clearly mention it.	2022-05-20 18:39:43 +02:00
Willy Tarreau	2b049b8166	CLEANUP: config: improve address parser error report for unmatched protocols Just trying "quic4@:4433" with USE_QUIC not set rsults in such a cryptic error: [ALERT] (14610) : config : parsing [quic-mini.cfg:44] : 'bind' : unsupported protocol family 2 for address 'quic4@:4433' Let's at least add the stream and datagram statuses to indicate what was being looked for: [ALERT] (15252) : config : parsing [quic-mini.cfg:44] : 'bind' : unsupported stream protocol for datagram family 2 address 'quic4@:4433' Still not very pretty but gives a little bit more info.	2022-05-20 18:39:43 +02:00
Willy Tarreau	0d04410ebe	BUG/MINOR: peers: fix error reporting of "bind" lines In case the str2listener() parser reports a generic error with no message when parsing the argument of a "bind" statement in a "peers" section, the reported error indicates an invalid address on the empty arg. This has existed since 2.0 with commit `355b2033e` ("MINOR: cfgparse: SSL/TLS binding in "peers" sections."), so this must be backported till 2.0.	2022-05-20 18:39:43 +02:00
Amaury Denoyelle	cc3d7166f4	MINOR: mux-quic: close connection on error if different data at offset As specified by the RFC reception of different STREAM data for the same offset should be treated with a CONNECTION_CLOSE with error PROTOCOL_VIOLATION. Use ncbuf API to detect this case : if add operation fails with NCB_RET_DATA_REJ with add mode NCB_ADD_COMPARE.	2022-05-20 17:56:00 +02:00
Amaury Denoyelle	209404bff1	MINOR: mux-quic: emit STREAM_LIMIT_ERROR Send a CONNECTION_CLOSE on reception of a STREAM frame for a STREAM id exceeding the maximum value enforced. Only implemented for bidirectional streams for the moment.	2022-05-20 17:52:07 +02:00
Amaury Denoyelle	d46b0f52ae	MINOR: mux-quic: emit FLOW_CONTROL_ERROR Send a CONNECTION_CLOSE if the peer emits more data than authorized by our flow-control. This is implemented for both stream and connection level. Fields have been added in qcc/qcs structures to differentiate received offsets for limit enforcing with consumed offsets for sending of MAX_DATA/MAX_STREAM_DATA frames.	2022-05-20 17:47:09 +02:00
Amaury Denoyelle	9fab9fd7e5	MINOR: quic/mux-quic: define CONNECTION_CLOSE send API Define an API to easily set a CONNECTION_CLOSE. This will mainly be useful for the MUX when an error is detected which require to close the whole connection. On the MUX side, a new flag is added when a CONNECTION_CLOSE has been prepared. This will disable add future send operations.	2022-05-20 17:26:56 +02:00
Frédéric Lécaille	dfd1301035	MINOR: quic: Dynamic Retry implementation We rely on <conn_opening> stats counter and tune.quic.retry_threshold setting to dynamically start sending Retry packets. We continue to send such packets when "quic-force-retry" setting is set. The difference is when we receive tokens. We check them regardless of this setting because the Retry could have been dynamically started. We must also send Retry packets when we receive Initial packets without token if the dynamic Retry threshold was reached but only for connection which are not currently opening or in others words for Initial packets without connection already instantiated. Indeed, we must not send Retry packets for all Initial packets without token. For instance a client may have already sent an Initial packet without receiving Retry packet because the Retry feature was not started, then the Retry starts on exeeding the threshold value due to others connections, then finally our client decide to send another Initial packet (to ACK Initial CRYPTO data for instance). It does this without token. So, for this already existing connection we must not send a Retry packet.	2022-05-20 17:11:13 +02:00
Frédéric Lécaille	9286210aa8	MINOR: quic: Add tune.quic.retry-threshold keyword This QUIC specific keyword may be used to set the theshold, in number of connection openings, beyond which QUIC Retry feature will be automatically enabled. Its default value is 100.	2022-05-20 17:11:13 +02:00
Frédéric Lécaille	cbd59c7ab6	MINOR: quic: QUIC stats counters handling First commit to handle the QUIC stats counters. There is nothing special to say except perhaps for ->conn_openings which is a gauge to count the number of connection openings. It is incremented after having instantiated a quic_conn struct, then decremented when the handshake was successful (handshake completed state) or failed or when the connection timed out without reaching the handshake completed state.	2022-05-20 17:11:13 +02:00
Frédéric Lécaille	3fd92f69e0	BUG/MINOR: quic: Fix potential memory leak during QUIC connection allocations Move the code which finalizes the QUIC connections initialisations after having called qc_new_conn() into this function to benefit from its error handling to release the memory allocated for QUIC connections the initialization of which could not be finalized.	2022-05-20 17:11:13 +02:00
Frédéric Lécaille	a89659a752	MINOR: quic: Attach proxy QUIC stats counters to the QUIC connection Make usage of EXTRA_COUNTERS_GET() do to so from qc_new_conn().	2022-05-20 17:11:13 +02:00
Frédéric Lécaille	a58cafeb89	MINOR: quic_stats: Add a new stats module for QUIC This is a very minimalist frontend only stats module with only one gauge for the QUIC establishing connections count.	2022-05-20 17:11:13 +02:00
Frédéric Lécaille	6492e66e41	MINOR: quic: Move quic_lstnr_dgram_dispatch() out of xprt_quic.c Remove this function from xprt_quic.c which for now implements only "by thread attached to a connection" code.	2022-05-20 16:57:12 +02:00
Frédéric Lécaille	ad20a56971	MINOR: cfgparse: Update for "cluster-secret" keyword for QUIC Retry The QUIC Retry feature is disabled if no "cluster-secret" setting was set.	2022-05-20 16:57:12 +02:00
Frédéric Lécaille	3f3ff47998	MINOR: quic: Retry implementation Here is the format of a token: - format (1 byte) - ODCID (from 9 up 21 bytes) - creation timestamp (4 bytes) - salt (16 bytes) A format byte is required to distinguish the Retry token from others sent in NEW_TOKEN frames. The Retry token is ciphered after having derived a strong secret from the cluster secret and generated the AEAD AAD, as well as a 16 bytes long salt. This salt is added to the token. Obviously it is not ciphered. The format byte is not ciphered too. The AAD are built by quic_generate_retry_token_aad() which concatenates the version, the client SCID and the IP address and port. We had to implement quic_saddr_cpy() to copy the IP address and port to the AAD buffer. Only the Retry SCID is generated on our side to build a Retry packet, the others fields come from the first packet received by the client. It must reuse this Retry SCID in response to our Retry packet. So, we have not to store it on our side. Everything is offloaded to the client (stateless). quic_generate_retry_token() must be used to generate a Retry packet. It calls quic_pkt_encrypt() to cipher the token. quic_generate_retry_check() must be used to check the validity of a Retry token. It is able to decipher a token which arrives into an Initial packet in response to a Retry packet. It calls parse_retry_token() after having deciphered the token to store the ODCID into a local quic_cid struct variable. Finally this ODCID may be stored into the transport parameter thanks to qc_lstnr_params_init(). The Retry token lifetime is 10 seconds. This lifetime is also checked by quic_generate_retry_check(). If quic_generate_retry_check() fails, the received packet is dropped without anymore packet processing at this time.	2022-05-20 16:57:12 +02:00
Frédéric Lécaille	55367c8679	MINOR: quic_tls: Add quic_tls_decrypt2() implementation This function does exactly the same thing as quic_tls_decrypt(), except that it does reuse its input buffer as output buffer. This is needed to decrypt the Retry token without modifying the packet buffer which contains this token. Indeed, this would prevent us from decryption the packet itself as the token belong to the AEAD AAD for the packet.	2022-05-20 16:57:12 +02:00
Frédéric Lécaille	a9c5d8da58	MINOR: quic_tls: Add quic_tls_derive_retry_token_secret() This function must be used to derive strong secrets from a non pseudo-random secret (cluster-secret setting in our case) and an IV. First it call quic_hkdf_extract_and_expand() to do that for a temporary strong secret (tmpkey) then two calls to quic_hkdf_expand() reusing this strong temporary secret to derive the final strong secret and IV.	2022-05-20 16:57:12 +02:00
Willy Tarreau	8ec9c81ac4	BUG/MINOR: cfgparse: abort earlier in case of allocation error In issue #1563, Coverity reported a very interesting issue about a possible UAF in the config parser if the config file ends in with a very large line followed by an empty one and the large one causes an allocation failure. The issue essentially is that we try to go on with the next line in case of allocation error, while there's no point doing so. If we failed to allocate memory to read one config line, the same may happen on the next one, and blatantly dropping it while trying to parse what follows it. In the best case, subsequent errors will be incorrect due to this prior error (e.g. a large ACL definition with many patterns, followed by a reference of this ACL). Let's just immediately abort in such a condition where there's no recovery possible. This may be backported to all versions once the issue is confirmed to be addressed. Thanks to Ilya for the report.	2022-05-20 09:20:12 +02:00
Amaury Denoyelle	3dde0d86dd	MINOR: quic: detect EBADF on sendto() EBADF can be encountered during process termination after fd listener has been reset to -1.	2022-05-19 11:58:32 +02:00
Amaury Denoyelle	ad5df386d9	MINOR: quic: abort on unlisted errno on sendto() If an unlisted errno is reported, abort the process. If a crash is reported on this condition, we must determine if the error code is a bug, should interrupt emission on the fd or if we can retry the syscall.	2022-05-19 10:18:18 +02:00
Amaury Denoyelle	8fa666650f	BUG/MINOR: quic: break for error on sendto If sendto returns an error, we should not retry the call and break from the sending loop. An exception is made for EINTR which allows to retry immediately the syscall. This bug caused an infinite loop reproduced when the process is in the closing state by SIGUSR1 but there is still QUIC data emission left.	2022-05-19 10:18:18 +02:00
Christopher Faulet	c95eaefbfd	MEDIUM: check: Use the CS to handle subscriptions for read/write events Instead of using the health-check to subscribe to read/write events, we now rely on the conn-stream. Indeed, on the server side, the conn-stream's endpoint is a multiplexer. Thus it seems appropriate to handle subscriptions for read/write events the same way than for the streams. Of course, the I/O callback function is not the same. We use srv_chk_io_cb() instead of cs_conn_io_cb().	2022-05-19 10:12:38 +02:00
Christopher Faulet	361417f9b4	REORG: check: Rename and export I/O callback function event_srv_chk_io() function is renamed srv_chk_io_cb() to be consistant with the I/O callback function of connections. In addition, this function is exported. It will be required to use the conn-stream's subscriptions.	2022-05-19 10:12:38 +02:00
Christopher Faulet	08c8f8e20d	MEDIUM: check: No longer shutdown the connection in .wake callback function The connection is already closed by the health-check itself. Thus there is now reason to duplicate this part in the .wake callback function. It is enough to wake the health-check and wait.	2022-05-19 10:12:38 +02:00
Christopher Faulet	6d781f612a	BUG/MINOR: check: Reinit the buffer wait list at the end of a check The buffer wait list is used to deal with buffer allocation failure. But at the end of health-check, it must be reinitialized. There is no reason to reason to get a buffer between two health-check runs. And in fact, the associated flags, CHK_ST_IN_ALLOC and CHK_ST_OUT_ALLOC, are already cleared at the end of a health-check. This patch must be backported as far as 2.2. On the 2.2, MT_LIST_ADDED and MT_LIST_DEL must be used instead of LIST_INLIST and LIST_DEL_INIT.	2022-05-19 10:12:38 +02:00
Christopher Faulet	dfe32c7e15	BUG/MEDIUM: config: Reset outline buffer size on realloc error in readcfgfile() When the line parsing failed because outline buffer must be reallocated, if my_realloc2() call fails, the buffer size must be reset. Indeed, in this case the current line is skipped, a fatal error is reported and we jump to the next line. At this stage the outline buffer is NULL. If the buffer size is not reset, the next call to parse_line() crashes because we try to write in the buffer. We fail to detect the outline buffer is too small to copy any character. To fix the issue, outlinesize variable must be set to 0 when outline allocation failed. This patch should fix the issue #1563. It must be backported as far as 2.2.	2022-05-19 10:12:38 +02:00
Amaury Denoyelle	8b87b15c22	MINOR: mux-quic: free RX buf if empty Release the QCS RX buffer if emptied afer qcs_consume(). This improves memory usage and avoids a QCS to keep an allocated buffer, particularly when no data is received anymore. Buffer is automatically reallocated if needed via qc_get_ncbuf().	2022-05-18 16:25:07 +02:00
Amaury Denoyelle	7313f5ee7e	BUG/MINOR: mux-quic: support nul buffer with qc_free_ncbuf() qc_free_ncbuf() may now be used with a NCBUF_NULL buffer as parameter. This is useful when using this function on a QCS with no allocated buffer. This case was not reproduced for the moment, but it will soon become more present as buffers will be released if emptied. Also a call to offer_buffers() is added to conform with the dynamic buffer management of haproxy.	2022-05-18 16:25:07 +02:00
Amaury Denoyelle	c830e1e904	MINOR: mux-quic: implement MAX_DATA emission This commit is similar to the previous one but deals with MAX_DATA for connection-level data flow control. It uses the same function qcc_consume_qcs() to update flow control level and generate a MAX_DATA frame if needed.	2022-05-18 16:25:07 +02:00
Amaury Denoyelle	a977355aa1	MINOR: mux-quic: implement MAX_STREAM_DATA emission Send MAX_STREAM_DATA frames when at least half of the allocated flow-control has been demuxed, frame and cleared. This is necessary to support QUIC STREAM with received data greater than a buffer. Transcoders must use the new function qcc_consume_qcs() to empty the QCS buffer. This will allow to monitor current flow-control level and generate a MAX_STREAM_DATA frame if required. This frame will be emitted via qc_io_cb().	2022-05-18 16:25:07 +02:00
Amaury Denoyelle	c985cb167d	MINOR: mux-quic: reorganize flow-control frames emission Adjust the mechanism for MAX_STREAMS_BIDI emission. When a bidirectional stream is removed, current flow-control level is checked. If needed, a MAX_STREAMS_BIDI frame is generated and inserted in a new list in the QCS instance. The new frames will be emitted at the start of qc_send(). This has no impact on the current MAX_STREAMS_BIDI behavior. However, this mechanism is more flexible and will allow to implement quickly MAX_STREAM_DATA/MAX_DATA emission.	2022-05-18 15:52:44 +02:00
Amaury Denoyelle	3a0864067a	MINOR: mux-quic: remove qcc_decode_qcs() call in XPRT Slightly change the interface for qcc_recv() between MUX and XPRT. The MUX is now responsible to call qcc_decode_qcs(). This is cleaner as now the XPRT does not have to deal with an extra QCS parameter and the MUX will call qcc_decode_qcs() only if really needed. This change is possible since there is no extra buffering for out-of-order STREAM frames and the XPRT does not have to handle buffered frames.	2022-05-18 15:50:57 +02:00
Amaury Denoyelle	37c2e4a65a	MEDIUM: mux-quic: implement recv on io-cb Previously, qc_io_cb() of mux-quic only dealt with TX. Add support for RX in it. This is done through a new function qc_recv(qcc). It loops over all QCS instances and call qcc_decode_qcs(qcs). This has no impact from the quic-conn layer as qcc_decode_qcs(qcs) is called directly. However, this allows to have a resume point when demux is blocked on the upper layer HTX full buffer. Note that for the moment, only RX for bidirectional streams is managed in qc_io_cb(). Unidirectional streams use their own mechanism for both TX/RX. It should be unified in the near future in a refactoring.	2022-05-18 15:43:22 +02:00
Amaury Denoyelle	73d6ffe832	MINOR: h3: flag demux as full on HTX full Flag QCS if HTX buffer is full on demux. This will block all future operations on QCS demux and should limit unnecessary decode_qcs() calls. The flag is cleared on rcv_buf operation called by conn-stream.	2022-05-18 15:38:01 +02:00
Amaury Denoyelle	b5454d42df	MINOR: h3: do not wait a complete frame for demuxing Previously, H3 demuxer refused to proceed the payload if the frame was not entirely received and the QCS buffer is not full. This code was duplicated from the H2 demuxer. In H2, this is a justified optimization as only one frame at a time can be demuxed. However, this is not the case in H3 with interleaved frames in the lower layer QUIC STREAM frames. This condition is now removed. H3 demuxer will proceed payload as soon as possible. An exception is kept for HEADERS frame as the code is not able to deal with partial HEADERS. With this change, H3 demuxer should consume less memory. To ensure that we never received a HEADER bigger than the RX buffer, we should use the H3 SETTINGS_MAX_FIELD_SECTION_SIZE.	2022-05-18 15:31:46 +02:00
Amaury Denoyelle	ca21c768b9	MINOR: ncbuf: refactor ncb_advance() First adjusted some typos in comments inside the function. Second, change the naming of some variable to reduce confusion. A special case has been inserted when advance is done inside a GAP block and this block is the last of the buffer. In this case, the whole buffer will be emptied, equivalent to a ncb_init() operation.	2022-05-18 15:30:13 +02:00
Amaury Denoyelle	f6dbdc1444	BUG/MINOR: ncbuf: fix ncb_is_empty() ncb_is_empty() was plainly incorrect as it directly dereferences the memory to read offset blocks instead of ncb_read_off(). The result is undefined. Also, BUG_ON() statement is wrong when the buffer starts with a data block. In this case, ncb_head() is not the first gap offset but instead just random data. The calculated sum in BUG_ON() statement has thus no meaning and may cause an abort. Adjust this by reorganizing the whole function. Only the first data block size is read. If and only if not nul, the first gap size is then checked. ncb_is_full() has been rewritten to share the same model as ncb_is_empty().	2022-05-18 15:23:29 +02:00
Amaury Denoyelle	80d0572a31	BUG/MEDIUM: quic: fix Rx buffering The quic-conn manages a buffer to store received QUIC packets. When the buffer wraps, the gap is filled until the end with junk and packets can be inserted at the start of the buffer. On the other end, deletion is implemented via quic_rx_pkts_del(). Packets are removed one by one if their refcount is nul. If junk is found, the buffer is emptied until its wrap. This seems to work in most cases but a bug was found in a particular case : on insertion if buffer gap is not at the end of the buffer. In this case, the gap was filled, which is useless as now the buffer is full and the packet cannot be inserted. Worst, on deletion, when junk is removed there is a risk to removed new packets. This can happens in the following case : 1. buffer contig space is too small, junk is inserted in the middle of it 2. on quic_rx_pkts_del() invocation, a packet is removed, but not the next one because its refcount is still positive. When a new packet is received, it will be stored after the junk. 3. on next quic_rx_pkts_del(), when junk is removed, all contig data is cleared, with newer packets data too. This will cause a transfer between a client and haproxy to be stalled. This can be reproduced with big enough POST requests. I triggered it with ngtcp2 and 10M of posted data. Hopefully, the solution of this bug is simple. If contig space is not big enough to store a packet, but the space is not at the end of the buffer, no junk is inserted and the packet is dropped as we cannot buffered it. This ensures that junk is only present at the end of the buffer and when removed no packets data is purged with it.	2022-05-18 15:02:14 +02:00
Christopher Faulet	f229e1894c	CLEANUP: httpclient: Remove useless test on ss_dst in httpclient_applet_init() In httpclient_applet_init() function, ss_dst variable is always defined before the call to sockaddr_alloc(). There is no reason to test it. This patch should fix the issue #1706.	2022-05-18 09:29:33 +02:00
Christopher Faulet	9e3c8d5512	CLEANUP: peers: Remove unreachable code in peer_session_create() An error label is now unreachable in peer_session_create(). This patch should fix the issue #1704.	2022-05-18 09:04:53 +02:00
Christopher Faulet	fa463afa8f	BUG/MINOR: spoe: Fix error handling in spoe_init_appctx() labels used in goto statement was not called in the right order. Thus if there is an error during the appctx startup, it is possible to leak a task. This patch should fix the issue #1703. No backport needed.	2022-05-18 09:04:53 +02:00
Tim Duesterhus	7ad27d41b4	CLEANUP: http_ana: Make use of the return value of stream_generate_unique_id() Even if `unique_id` and `s->unique_id` are identical it is a bit odd to `isttest()` `unique_id` and then use `s->unique_id` in the call to `http_add_header()`. This "issue" was introduced in `a17e66289c`, because before that commit the function returned the length of the ID, as it was not an ist.	2022-05-18 07:19:01 +02:00
Remi Tricot-Le Breton	ccc0355c41	MINOR: ssl: Add 'ssl-provider-path' global option When loading providers with 'ssl-provider' global options, this ssl-provider-path option can be used to set the search path that is to be used by openssl. It behaves the same way as the OPENSSL_MODULES environment variable.	2022-05-17 18:09:17 +02:00
Christopher Faulet	bc684acae7	CLEANUP: proxy: Remove dead code when parsing "http-restrict-req-hdr-names" option negation or default modifiers are not supported for this option. However, this was already tested earlier in cfg_parse_listen() function. Thus, when "http-restrict-req-hdr-names" option is parsed, the keyword modifier is always equal to KWM_STD. It is useless to test it again at this place. This patch should solve the issue #1702.	2022-05-17 16:13:22 +02:00
Christopher Faulet	2d9cc85b74	MINOR: conn-stream/applet: Stop setting appctx as the endpoint context The appctx is already the endpoint target. It is confusing to also use it to set the endpoint context. So, never set the endpoint ctx when an appctx is created or attached to an existing conn-stream.	2022-05-17 16:13:22 +02:00
Maciej Zdeb	34e4085f8a	MEDIUM: peers: Balance applets across threads When creating a new applet for peer outgoing connection, we check the load on each thread. Threads with least applet count are preferred. With this solution we avoid a situation when many outgoing connections run on the same thread causing significant load on single CPU core.	2022-05-17 16:13:22 +02:00
Maciej Zdeb	d01be2ab13	MINOR: peers: Track number of applets run by thread Maintain number of peers applets run on all threads. It will be used in next patch for least loaded thread selection.	2022-05-17 16:13:22 +02:00
Christopher Faulet	d9c1d33fa1	MEDIUM: applet: Add support for async appctx startup on a thread subset It is now possible to start an appctx on a thread subset. Some controls were added here and there. It is forbidden to start a backend appctx on another thread than the local one. If a frontend appctx is started on another thread or a thread subset, the applet .init callback function must be defined. This callback function is responsible to finalize the appctx startup. It can be performed synchornously. In this case, the appctx is started on the local thread. It is not really useful but it is valid. Or it can be performed asynchronously. In this case, .init callback function is called when the appctx is woken up for the first time. When this happens, the appctx affinity is set to the current thread to be able to start the session and the stream.	2022-05-17 16:13:22 +02:00
Christopher Faulet	6095d57701	MINOR: applet: Add API to start applet on a thread subset In the same way than for the tasks, the applets api was changed to be able to start a new appctx on a thread subset. For now the feature is disabled. Only appctx_new_here() is working. But it will be possible to start an appctx on a specific thread or a subset via a mask.	2022-05-17 16:13:22 +02:00
Christopher Faulet	6712dc680c	MEDIUM: peers: Refactor peer appctx creation A .init callback function is defined for the peer_applet applet. This function finishes the appctx startup by calling appctx_finalize_startup() and its handles the stream customization.	2022-05-17 16:13:22 +02:00
Christopher Faulet	387e79727c	MINOR: peers: Add a ref to peers section in the peer structure This change is required to handle asynchrone init of the appctx. It is now possible to directly get the peers section associated to a peer.	2022-05-17 16:13:22 +02:00
Christopher Faulet	aee57fc277	MEDIUM: sink: Refactor sink forwarder appctx creation A .init callback function is defined for the sink_forward_applet applet. This function finishes the appctx startup by calling appctx_finalize_startup() and its handles the stream customization.	2022-05-17 16:13:22 +02:00
Christopher Faulet	2ae25ea24b	MINOR: sink: Add a ref to sink in the sink_forward_target structure This change is required to be able to refactor the init stage of appctx. It is now possible to directly get the sink from a forward target.	2022-05-17 16:13:22 +02:00
Christopher Faulet	b1e0836ecf	MEDIUM: httpclient: Refactor http-client appctx creation A .init callback function is defined for the httpclient_applet applet. This function finishes the appctx startup by calling appctx_finalize_startup() and its handles the stream customization.	2022-05-17 16:13:22 +02:00
Christopher Faulet	5f35a3e2ec	MEDIUM: lua: Refactor cosocket appctx creation A .init callback function is defined for the update_applet applet. This function finishes the appctx startup by calling appctx_finalize_startup() and its handles the stream customization.	2022-05-17 16:13:22 +02:00
Christopher Faulet	69ebc3093a	MEDIUM: spoe: Refactor SPOE appctx creation A .init callback function is defined for the spoe_applet applet. This function finishes the spoe_appctx initialization. It also finishes the appctx startup by calling appctx_finalize_startup() and its handles the stream customization.	2022-05-17 16:13:22 +02:00
Christopher Faulet	9223851a80	MEDIUM: dns: Refactor dns appctx creation A .init callback function is defined for the dns_session_applet applet. This function finishes the appctx startup by calling appctx_finalize_startup() and its handles the stream customization.	2022-05-17 16:13:21 +02:00
Christopher Faulet	d0c4ec04b8	MINOR: applet: Add function to release appctx on error during init stage appctx_free_on_early_error() must be used to release a freshly created frontend appctx if an error occurred during the init stage. It takes care to release the stream instead of the appctx if it exists. For a backend appctx, it just calls appctx_free().	2022-05-17 16:13:21 +02:00
Christopher Faulet	8718c95c0a	MINOR: applet: Add a function to finalize frontend appctx startup appctx_finalize_startup() may be used to finalize the frontend appctx startup. It is responsible to create the appctx's session and the frontend conn-stream. On error, it is the caller responsibility to release the appctx. However, the session is released if it was created. On success, if an error is encountered in the caller function, the stream must be released instead of the appctx. This function should ease the init stage when new appctx is created.	2022-05-17 16:13:21 +02:00
Christopher Faulet	16c0d9cda0	MINOR: applet: Add appctx_init() helper fnuction It is just a helper function that call the .init applet callback function, if it exists. This will simplify a bit the init stage when a new applet is started. For now, this callback function is only used when a new service is started.	2022-05-17 16:13:21 +02:00
Christopher Faulet	ab5d1dceed	MINOR: stream: Export stream_free() The stream_free() function is now public. It is mandatory to properly handle errors when a new applet is started.	2022-05-17 16:13:21 +02:00
Christopher Faulet	c9929380a4	MINOR: applet: Change return value for .init callback function 0 is now returned on success and -1 on error.	2022-05-17 16:13:21 +02:00
Christopher Faulet	92202da2da	MINOR: applet: Let the frontend appctx release the session The session created for frontend applets is now totally owns by the corresponding appctx. It means the appctx is now responsible to release it. This removes the hack in stream_free() about frontend applets to be sure to release the session.	2022-05-17 16:13:21 +02:00
Christopher Faulet	ac57bb527a	MINOR: applet: Prepare appctx to own the session on frontend side Applets were moved at the same level than multiplexers. Thus, gradually, applets code is changed to be less dependent from the stream. With this commit, the frontend appctx are ready to own the session. It means a frontend appctx will be responsible to release the session.	2022-05-17 16:13:21 +02:00
Remi Tricot-Le Breton	9bf3a1f67e	BUG/MINOR: ssl: Fix crash when no private key is found in pem If no private key can be found in a bind line's certificate and ssl-load-extra-files is set to none we end up trying to call X509_check_private_key with a NULL key, which crashes. This fix should be backported to all stable branches.	2022-05-17 15:51:41 +02:00
David Carlier	7198c700bc	MINOR: tools: add get_exec_path implementation for solaris based systems. We can use getexecname() which fetches AT_SUN_EXECNAME from the auxiliary vectors.	2022-05-17 11:44:21 +02:00
Tim Duesterhus	9b6becb312	CLEANUP: Remove unused function hlua_get_top_error_string This function has no prototype defined in a header and is not used in hlua.c either, thus it can be safely removed. Found with -Wmissing-prototypes.	2022-05-17 11:40:33 +02:00
Tim Duesterhus	6eded62f6a	CLEANUP: Add missing header to hlua_fcn.c Found with -Wmissing-prototypes: src/hlua_fcn.c:53:5: fatal error: no previous prototype for function 'hlua_checkboolean' [-Wmissing-prototypes] int hlua_checkboolean(lua_State L, int index) ^ src/hlua_fcn.c:53:1: note: declare 'static' if the function is not intended to be used outside of this translation unit int hlua_checkboolean(lua_State L, int index) ^ static 1 error generated.	2022-05-17 11:40:33 +02:00
Tim Duesterhus	2a0688aa89	CLEANUP: Add missing header to ssl_utils.c Found with -Wmissing-prototypes: src/ssl_utils.c:22:5: fatal error: no previous prototype for function 'cert_get_pkey_algo' [-Wmissing-prototypes] int cert_get_pkey_algo(X509 crt, struct buffer out) ^ src/ssl_utils.c:22:1: note: declare 'static' if the function is not intended to be used outside of this translation unit int cert_get_pkey_algo(X509 crt, struct buffer out) ^ static 1 error generated.	2022-05-17 11:40:33 +02:00
Remi Tricot-Le Breton	1746a388c5	MINOR: ssl: Add 'ssl-provider' global option When HAProxy is linked to an OpenSSLv3 library, this option can be used to load a provider during init. You can specify multiple ssl-provider options, which will be loaded in the order they appear. This does not prevent OpenSSL from parsing its own configuration file in which some other providers might be specified. A linked list of the providers loaded from the configuration file is kept so that all those providers can be unloaded during cleanup. The providers loaded directly by OpenSSL will be freed by OpenSSL.	2022-05-17 10:56:05 +02:00
Remi Tricot-Le Breton	e80976526c	MINOR: ssl: Add 'ssl-propquery' global option This option can be used to define a default property query used when fetching algorithms in OpenSSL providers. It follows the format described in https://www.openssl.org/docs/man3.0/man7/property.html. It is only available when haproxy is built with SSL support and linked to OpenSSLv3 libraries.	2022-05-17 10:56:05 +02:00
Remi Tricot-Le Breton	5194446b76	MEDIUM: ssl: Delay random generator initialization after config parsing The random generator initialization needs to be performed before the chroot but it is not needed before. If we want to add provider configuration option to the configuration file, they need to be processed before any call to a crypto-related OpenSSL function. We can then delay the initialization until after the configuration file is parsed and processed.	2022-05-17 10:55:59 +02:00
Christopher Faulet	18c13d3bd8	MEDIUM: http-ana: Add a proxy option to restrict chars in request header names The "http-restrict-req-hdr-names" option can now be set to restrict allowed characters in the request header names to the "[a-zA-Z0-9-]" charset. Idea of this option is to not send header names with non-alphanumeric or hyphen character. It is especially important for FastCGI application because all those characters are converted to underscore. For instance, "X-Forwarded-For" and "X_Forwarded_For" are both converted to "HTTP_X_FORWARDED_FOR". So, header names can be mixed up by FastCGI applications. And some HAProxy rules may be bypassed by mangling header names. In addition, some non-HTTP compliant servers may incorrectly handle requests when header names contain characters ouside the "[a-zA-Z0-9-]" charset. When this option is set, the policy must be specify: * preserve: It disables the filtering. It is the default mode for HTTP proxies with no FastCGI application configured. * delete: It removes request headers with a name containing a character outside the "[a-zA-Z0-9-]" charset. It is the default mode for HTTP backends with a configured FastCGI application. * reject: It rejects the request with a 403-Forbidden response if it contains a header name with a character outside the "[a-zA-Z0-9-]" charset. The option is evaluated per-proxy and after http-request rules evaluation. This patch may be backported to avoid any secuirty issue with FastCGI application (so as far as 2.2).	2022-05-16 16:00:26 +02:00
Amaury Denoyelle	f46393ac44	MINOR: ncbuf: fix warnings for testing build Using -Wall reveals several warning when building ncbuf testing API. One of them was about the signedness mismatch. The other one was with an incorrect print format.	2022-05-16 11:32:33 +02:00
Amaury Denoyelle	48fbad45e2	BUG/MEDIUM: ncbuf: fix null buffer usage ncbuf public API functions were not ready to deal with a NCBUF_NULL as parameter. Strenghten these functions by handling it properly. Most of the functions will consider the buffer as empty and silently returns. The only exception is ncb_init(buf) which cannot be called with a NCBUF_NULL. This seems legitimate to consider this as a bug and not silently failed in this case.	2022-05-16 11:28:41 +02:00
Amaury Denoyelle	45fce8fcb5	CLEANUP: quic: remove unused quic_rx_strm_frm quic_rx_strm_frm type was used to buffered STREAM frames received out of order. Now the MUX is able to deal directly with these frames and buffered it inside its ncbuf.	2022-05-13 17:29:52 +02:00
Amaury Denoyelle	00f87bbaa3	CLEANUP: mux-quic: remove unused fields for Rx Rx has been simplified since the conversion of buffer to a ncbuf. The old buffer can now be removed. The frms tree is also removed. It was used previously to stored out-of-order received STREAM frames. Now the MUX is able to buffer them directly into the ncbuf.	2022-05-13 17:29:52 +02:00
Amaury Denoyelle	3db98e9d13	MEDIUM: mux-quic/h3/qpack: use ncbuf for uni streams This commit is the equivalent for uni-streams of previous commit MEDIUM: mux-quic/h3/hq-interop: use ncbuf for bidir streams All unidirectional streams data is now handle in MUX Rx ncbuf. The obsolete buffer is not unused and will be cleared in the following patches.	2022-05-13 17:29:49 +02:00
Amaury Denoyelle	1290f1ebfb	MEDIUM: mux-quic/h3/hq-interop: use ncbuf for bidir streams Add a ncbuf for data reception on qcs. Thanks to this, the MUX is able to buffered all received frame directly into the buffer. Flow control parameters will be used to ensure there is never an overflow. This change will simplify Rx path with the future deletion of acked frames tree previously used for frames out of order.	2022-05-13 17:28:46 +02:00
Amaury Denoyelle	b5ca943ff9	BUG/MINOR: ncbuf: fix coverity warning on uninit sz_data Coverity reports that data block generated by ncb_blk_first() has sz_data field uninitialized. This has no real impact as it has no sense for data block. Set to 0 to hide the warning. This should fix github issue #1695.	2022-05-13 17:06:24 +02:00
Willy Tarreau	26f221bd55	MINOR: config: make sure never to mix dgram and stream protocols on a bind line It is absolutely not possible to use the same "bind" line to listen to both quic and tcp for example, because no single transport layer would fit both modes and we'll need the type to choose one then to choose a mux. Let's make sure this does not happen. This may be relaxed in the future if we manage to instantiate transport layers on the fly, but the SSL vs quic part might be tricky to handle.	2022-05-13 17:00:09 +02:00
Willy Tarreau	8d31ab0438	MINOR: tools: improve error message accuracy in str2sa_range The error message when mixing stream and dgram protocols in an address speaks about sockets while it ought to speak about addresses, let's fix this as in some contexts it can be a bit confusing.	2022-05-13 16:50:00 +02:00
Willy Tarreau	df94313e0e	BUG/MEDIUM: mux-quic: fix a thinko in the latest cs/endpoint cleanup Fred & Amaury found that I messed up with qc_detach() in commit `4201ab791` ("CLEANUP: muxes: make mux->attach/detach take a conn_stream endpoint"), causing a segv in this case with endp->cs == NULL being passed to __cs_mux(). It obviously ought to have been endp->target like in other muxes. No backport needed.	2022-05-13 16:34:40 +02:00
Willy Tarreau	973cf90714	MINOR: ext-check: indicate the transport and protocol of a server Valerio Pachera explained [1] that external checks would benefit from having a variable indicating if SSL is being used or not on the server being checked, and the discussion derived to also indicating the protocol in use. This patch adds two environment variables for external checks: - HAPROXY_SERVER_SSL: equals "0" when SSL is not used, "1" when it is - HAPROXY_SERVER_PROTO: contains one of the following words to describe the protocol used with this server: - "cli": the haproxy CLI. Normally not seen - "syslog": this is a syslog TCP server - "peers": this is a peers TCP server - "h1": this is an HTTP/1.x server - "h2": this is an HTTP/2 server - "tcp": this is any other TCP server The patch is very simple, and may be backported to recent versions if needed. This closes github issue #1692. [1] https://www.mail-archive.com/haproxy@formilux.org/msg42233.html	2022-05-13 16:06:29 +02:00
Willy Tarreau	6796a06278	CLEANUP: conn_stream: merge cs_new_from_{mux,applet} into cs_new_from_endp() The two functions became exact copies since there's no more special case for the appctx owner. Let's merge them into a single one, that simplifies the code.	2022-05-13 14:28:48 +02:00
Willy Tarreau	0698c80a58	CLEANUP: applet: remove the unneeded appctx->owner This one is the pointer to the conn_stream which is always in the endpoint that is always present in the appctx, thus it's not needed. This patch removes it and replaces it with appctx_cs() instead. A few occurences that were using __cs_strm(appctx->owner) were moved directly to appctx_strm() which does the equivalent.	2022-05-13 14:28:48 +02:00
Willy Tarreau	1c3ead45a4	MINOR: applet: replace cs_applet_shut() with appctx_shut() The former takes a conn_stream still attached to a valid appctx, which also complicates the termination of the applet. Instead, let's pass the appctx which already points to the endpoint, this allows us to properly detach the conn_stream before the call, which is cleaner and safer.	2022-05-13 14:28:48 +02:00
Willy Tarreau	4201ab791d	CLEANUP: muxes: make mux->attach/detach take a conn_stream endpoint The mux ->detach() function currently takes a conn_stream. This causes an awkward situation where the caller cs_detach_endp() has to partially mark it as released but not completely so that ->detach() finds its endpoint and context, and it cannot be done later since it's possible that ->detach() deletes the endpoint. As such the endpoint link between the conn_stream and the mux's stream is in a transient situation while we'd like it to be clean so that the mux's ->detach() code can call any regular function it wants that knows the regular semantics of the relation between the CS and the endpoint. A better approach consists in slightly modifying the detach() API to better match the reality, which is that the endpoint is detached but still alive and that it's the only part the function is interested in. As such, this patch modifies the function to take an endpoint there, and by analogy (or simplicity) does the same for ->attach(), even though it looks less important there since we're always attaching an endpoint to a conn_stream anyway. It is possible that in the future the API could evolve to use more endpoints that provide a bit more flexibility in the API, but at this point we don't need to go further.	2022-05-13 14:28:48 +02:00
Willy Tarreau	cfbfc3f091	MINOR: mux-pt: remove the now unneeded conn_stream from the context Since we always have a valid endpoint we can safely use it to access the conn_stream and stop using ctx->cs. That's one less pointer to care about.	2022-05-13 14:28:48 +02:00
Willy Tarreau	01c2a4a86f	MINOR: mux-quic: remove the now unneeded conn_stream from the qcs Since we always have a valid endpoint we can safely use it to access the conn_stream and stop using qcs->cs. That's one less pointer to care about.	2022-05-13 14:28:48 +02:00
Willy Tarreau	b57669e6a4	MINOR: mux-fcgi: remove the now unneeded conn_stream from the fcgi_strm Since we always have a valid endpoint we can safely use it to access the conn_stream and stop using fstrm->cs. That's one less pointer to care about.	2022-05-13 14:28:48 +02:00
Willy Tarreau	c84610c2ac	MINOR: mux-fcgi: make sure any stream always has an endpoint The principle that each mux stream should have an endpoint is not guaranteed for closed streams that map to the dummy static streams. Let's have a dummy endpoint for use with such streams. It only has the DETACHED flag and a NULL conn_stream, and is referenced by all the closed streams so that we can afford not to test strm->endp when trying to access the flags or the CS.	2022-05-13 14:28:48 +02:00
Willy Tarreau	cd6bb1a5e5	MINOR: mux-h2: remove the now unneeded conn_stream from the h2s Since we always have a valid endpoint we can safely use it to access the conn_stream and stop using h2s->cs. That's one less pointer to care about.	2022-05-13 14:28:48 +02:00
Willy Tarreau	b22b5f02af	MINOR: mux-h2: make sure any h2s always has an endpoint The principle that each mux stream should have an endpoint is not guaranteed for closed streams that map to the dummy static streams. Let's have a dummy endpoint for use with such streams. It only has the DETACHED flag and a NULL conn_stream, and is referenced by all the closed streams so that we can afford not to test h2s->endp when trying to access the flags or the CS.	2022-05-13 14:28:48 +02:00
Willy Tarreau	56d5a819d7	MINOR: mux-h1: remove the now unneeded h1s->cs There is always an endpoint link in a stream, and this endpoint link contains a pointer to the conn_stream it's attached to, so the one in the h1 stream is always duplicate now. Let's always use endp->cs instead and get rid of it.	2022-05-13 14:28:48 +02:00
Willy Tarreau	efb4618c6e	MINOR: conn_stream: add a pointer back to the cs from the endpoint Muxes and applets need to have both a pointer to the endpoint and to the conn_stream. It would seem more natural that they only have a pointer to the endpoint (that is always there) and that this one has an optional pointer to the conn_stream. This would reduce the number of elements to manipulate in lower level code. In addition, the conn_stream is not much used from the lower layers (wake and exceptional events mostly).	2022-05-13 14:28:48 +02:00
Willy Tarreau	66435e5f63	CLEANUP: applet: use the appctx's endp instead of cs->endp The few applets that set CS_EP_EOI or CS_EP_ERROR used to set it on the endpoint retrieved from the conn_stream while it's already available on the appctx itself. Better use the appctx one to limit the unneeded interactions between the two sides.	2022-05-13 14:28:46 +02:00
Willy Tarreau	15b0721ca5	CLEANUP: mux-quic: always take the endp from the qcs not the cs At a few places the endpoint pointer was retrieved from the conn_stream while it's safer and more long-term proof to take it from the qcs. Let's just do that.	2022-05-13 14:27:57 +02:00
Willy Tarreau	7d299c284b	CLEANUP: mux-fcgi: always take the endp from the fstrm not the cs At a few places the endpoint pointer was retrieved from the conn_stream while it's safer and more long-term proof to take it from the fstrm. Let's just do that.	2022-05-13 14:27:57 +02:00
Willy Tarreau	7a2705f921	CLEANUP: mux-pt: always take the endp from the context not the cs At a few places the endpoint pointer was retrieved from the conn_stream while it's safer and more long-term proof to take it from the context. Let's just do that.	2022-05-13 14:27:57 +02:00
Willy Tarreau	aff21f9a96	CLEANUP: mux-h2: always take the endp from the h2s not the cs At a few places the endpoint pointer was retrieved from the conn_stream while it's safer and more long-term proof to take it from the h2s. Let's just do that.	2022-05-13 14:27:57 +02:00
Willy Tarreau	61533d3486	CLEANUP: mux-h1: always take the endp from the h1s not the cs At a few places the endpoint pointer was retrieved from the conn_stream while it's safer and more long-term proof to take it from the h1s. Let's just do that.	2022-05-13 14:27:57 +02:00
Willy Tarreau	386346f5eb	MINOR: conn_stream: make cs_set_error() work on the endpoint instead Wherever we need to report an error, we have an even easier access to the endpoint than the conn_stream. Let's first adjust the API to use the endpoint and rename the function accordingly to cs_ep_set_error().	2022-05-13 14:27:57 +02:00
Christopher Faulet	b112b1d02a	CLEANUP: mux-h1: Fix comments and error messages for global options Wrong name was used in comments and error messages for "h1-header-case-adjust" and "h1-headers-case-adjust-file" global options.	2022-05-13 12:04:24 +02:00
Christopher Faulet	0f9c0f5801	MINOR: mux-h1: Add global option accpet payload for any HTTP/1.0 requests Since the 2.5, for security reason, HTTP/1.0 GET/HEAD/DELETE requests with a payload are rejected (See `e136bd12a` "MEDIUM: mux-h1: Reject HTTP/1.0 GET/HEAD/DELETE requests with a payload" for details). However it may be an issue for old clients. To avoid any compatibility issue with such clients, "h1-accept-payload-with-any-method" global option was added. It must only be set if there is a good reason to do so because it may lead to a request smuggling attack on some servers or intermediaries. This patch should solve the issue #1691. it may be backported to 2.5.	2022-05-13 12:04:24 +02:00
William Lallemand	ae053b30da	BUG/MEDIUM: wdt: don't trigger the watchdog when p is unitialized In wdt_handler(), does not try to trigger the watchdog if the prev_cpu_time wasn't initialized. This prevents an unexpected trigger of the watchdog when it wasn't initialized yet. This case could happen in the master just after loading the configuration. This would show a trace where the <diff> value is equal to the <now> value in the trace, and the <poll> value would be 0. For example: Thread 1 is about to kill the process. *>Thread 1 : id=0x0 act=1 glob=1 wq=0 rq=0 tl=0 tlsz=0 rqsz=0 stuck=1 prof=0 harmless=0 wantrdv=0 cpu_ns: poll=0 now=6005541706 diff=6005541706 curr_task=0 Thanks to Christian Ruppert for repporting the problem. Could be backported in every stable versions.	2022-05-13 11:28:08 +02:00
Boyang Li	e0c5435d76	BUG/MEDIUM: lua: fix argument handling in data removal functions Lua API Channel.remove() and HTTPMessage.remove() expects 1 to 3 arguments (counting the manipulated object), with offset and length being the 2nd and 3rd argument, respectively. hlua_{channel,http_msg}_del_data() incorrectly gets the 3rd argument as offset, and 4th (nonexistent) as length. hlua_http_msg_del_data() also improperly checks arguments. This patch fixes argument handling in both. Must be backported to 2.5.	2022-05-13 08:40:03 +02:00
Amaury Denoyelle	eeeeed44ea	MINOR: ncbuf: write unit tests Implement a series of unit test to validate ncbuf. This is written with a main function which can be compiled independently using the following command-line : $ gcc -DSTANDALONE -lasan -I./include -o ncbuf src/ncbuf.c The first part tests is used to test ncb_add()/ncb_advance(). After each call a loop is done on the buffer blocks which should ensure that the gap infos are correct. The second part generates random offsets and insert them until the buffer is full. The buffer is then resetted and all random offsets are re-inserted in the reverse order : the buffer should be full once again. The generated binary takes arguments to change the tests execution. "usage: ncbuf [-r] [-s bufsize] [-h bufhead] [-p <delay_msec>]"	2022-05-12 18:29:55 +02:00
Amaury Denoyelle	df25acf47f	MINOR: ncbuf: implement advance A new function ncb_advance() is implemented. This is used to advance the buffer head pointer. This will consume the front data while forming a new gap at the end for future data. On success NCB_RET_OK is returned. The operation can be rejected if a too small new gap is formed in front of the buffer.	2022-05-12 18:29:55 +02:00
Amaury Denoyelle	b830f0d8d9	MINOR: ncbuf: define various insertion modes Define three different ways to proceed insertion. This configures how overlapping data is treated. - NCB_ADD_PRESERVE : in this mode, old data are kept during insertion. - NCB_ADD_OVERWRT : new data will overwrite old ones. - NCB_ADD_COMPARE : this mode adds a new test in check stage. The overlapping old and new data must be identical or else the insertion is not conducted. An error NCB_RET_DATA_REJ is used in this case. The mode is specified with a new argument to ncb_add() function.	2022-05-12 18:27:05 +02:00
Amaury Denoyelle	077e096b30	MINOR: ncbuf: implement insertion Implement a new function ncb_add() to insert data in ncbuf. This operation is conducted in two stages. First, a simulation will be run to ensure that insertion can be proceeded. If a gap is formed, either before or after the new data, it must be big enough to store its header, or else the insertion is aborted. After this check stage, the insertion is conducted block by block with the function pair ncb_fill_data_blk()/ncb_fill_gap_blk(). A new type ncb_ret is used as a return value. For the moment, only success or gap-size error is used. It is planned to add new error types in the future when insertion will be extended.	2022-05-12 18:27:05 +02:00
Amaury Denoyelle	edeb0a61a2	MINOR: ncbuf: optimize storage for the last gap Relax the constraint for gap storage when this is the last block. ncb_blk API functions will consider that if a gap is stored near the end of the buffer, without the space to store its header, the gap will cover entirely the buffer end. For these special cases, the gap size/data size are not write/read inside the gap to prevent an overflow. Such a gap is designed in functions as "reduced gap" and will be flagged with the value NCB_BK_F_FIN. This should reduce the rejection on future add operation when receiving data in-order. Without reduced gap handling, an insertion would be rejected if it covers only partially the last buffer bytes, which can be a very common case.	2022-05-12 18:18:47 +02:00
Amaury Denoyelle	d5d2ed90f0	MINOR: ncbuf: complete API and define block interal abstraction Implement two new functions to report the total data stored accross the whole buffer and the data stored at a specific offset until the next gap or the buffer end. To facilitate implementation of these new functions and also future add/delete operations, a new abstraction is introduced : ncb_blk. This structure represents a block of either data or gap in the buffer. It simplifies operation when moving forward in the buffer. The first buffer block can be retrieved via ncb_blk_first(buf). The block at a specific offset is accessed via ncb_blk_find(buf, off). This abstraction is purely used in functions but not stored in the ncbuf structure per-se. This is necessary to keep the minimal memory footprint.	2022-05-12 18:18:47 +02:00
Amaury Denoyelle	1b5f77fc18	MINOR: ncbuf: define non-contiguous buffer Define the new type ncbuf. It can be used as a buffer with non-contiguous data and wrapping support. To reduce as much as possible the memory footprint, size of data and gaps are stored in the gaps themselves. This put some limitation on the buffer usage. A reserved space is present just before the head to store the size of the first data block. Also, add and delete operations will be constrained to ensure minimal gap sizes are preserved. The sizes stored in the gaps are represented by a custom type named ncb_sz_t. This type is a typedef to easily change it : this has a direct impact on the maximum buffer size (MAX(ncb_sz_t) - sizeof(ncb_sz_t)) and the minimal gap sizes (sizeof(ncb_sz_t) * 2)). Currently, it is set to uint32_t.	2022-05-12 18:13:21 +02:00
Frédéric Lécaille	4ba3b4ef67	CLEANUP: quic: Useless use of pointer for quic_hkdf_extract() There is no need to use a pointer to the output buffer length.	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	a54e49d0b1	CLEANUP: quic: wrong use of ebentry() macro This wrong use has no consequence because the ->node member fields of ebnode structs are the first.	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	36b28ed012	MINOR: quic: Short packets always embed a trailing AEAD TAG We must drop as soon as possible too small 1-RTT packets to be valid QUIC packets to avoid replying with stateless reset packets.	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	e2fb1bf487	MINOR: quic: Send stateless reset tokens Add send_stateless_reset() to send a stateless reset packet. It prepares a packet to build a 1-RTT packet with quic_stateless_reset_token_cpy() to copy a stateless reset token derived from the cluster secret with the destination connection ID received as salt. Also add QUIC_EV_STATELESS_RST new trace event to at least to have a trace of the connection which are reset.	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	806e6cf392	MINOR: quic: Stateless reset token copy to transport parameters A server may send the stateless reset token associated to the current connection from its transport parameters. So, let's copy it from qc_lstnt_params_init().	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	395a64dd81	MINOR: qc_new_conn() rework for stateless reset The stateless reset token of a connection is generated from qc_new_conn() when allocating the first connection ID. A QUIC server can copy it into its transport parameters to allow the peer to reset the associated connection. This latter is not easily reachable after having returned from qc_new_conn(). We want to be able to initialize the transport parameters from this function which has an access to all the information to do so. Extract the code used to initialize the transport parameters from qc_lstnr_pkt_rcv() and make it callable from qc_new_conn(). qc_lstnr_params_init() is implemented to accomplish this task for a haproxy listener. Modify qc_new_conn() to reduce its the number of parameters. The source address coming from Initial packets is also copied from qc_new_conn().	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	28a1795515	MINOR: quic: Initialize stateless reset tokens with HKDF secrets Add quic_stateless_reset_token_init() wrapper function around quic_hkdf_extract_and_expand() function to derive the stateless reset tokens attached to the connection IDs from "cluster-secret" configuration setting and call it each time we instantiate a QUIC connection ID.	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	0226c521b0	MINOR: quic: new_quic_cid() code moving This function will have to call another one from quic_tls.[ch] soon. As we do not want to include quic_tls.h from xprt_quic.h because quic_tls.h already includes xprt_quic.h, let's moving it into xprt_quic.c.	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	7b92c81e43	MINOR: quic-tls: Add quic_hkdf_extract_and_expand() for HKDF This is a wrapper function around OpenSSL HKDF API functions to use the "extract-then-expand" HKDF mode as defined by rfc5869. This function will be used to derived stateless reset tokens from secrets ("cluster-secret" conf. keyword) and CIDs (as salts).	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	372508cc42	MINOR: config: Add "cluster-secret" new global keyword It could be usefull to set a ASCII secret which could be used for different usages. For instance, it will be used to derive QUIC stateless reset tokens.	2022-05-12 17:48:35 +02:00
Frédéric Lécaille	7cc8b3166a	MINOR: quic: Add correct ack delay values to ACK frames A ->time_received new member is added to quic_rx_packet to store the time the packet are received. ->largest_time_received is added the the packet number space structure to store this timestamp for the packet with a new largest packet number to be acknowledged. QUIC_FL_PKTNS_NEW_LARGEST_PN new flag is added to mark a packet number space as having to acknowledged a packet wih a new largest packet number. In this case, the packet number space ack delay must be recalculated. Add quic_compute_ack_delay_us() function to compute the ack delay from the value of the time a packet was received. Used only when a packet with a new largest packet number.	2022-05-12 15:30:14 +02:00
Frédéric Lécaille	8726d633d4	MINOR: quic: Add a debug counter for sendto() errors As we do not have any task to be wake up by the poller after sendto() error, we add an sendto() error counter to the quic_conn struct. Dump its values from qc_send_ppkts().	2022-05-12 15:11:53 +02:00
Willy Tarreau	e872f75fc4	MINOR: mux-h2: report a trace event when failing to create a new stream There are two reasons we can reject the creation of an h2 stream on the frontend: - its creation would violate the MAX_CONCURRENT_STREAMS setting - there's no more memory available And on the backend it's almost the same except that the setting might have be negotiated after trying to set up the stream. Let's add traces for such a sitaution so that it's possible to know why the stream was rejected (currently we only know it was rejected). It could be nice to backport this to the most recent versions.	2022-05-12 09:29:58 +02:00
Willy Tarreau	198b50770d	BUG/MINOR: mux-h2: mark the stream as open before processing it not after When a client doesn't respect the h2 MAX_CONCURRENT_STREAMS setting, we rightfully send RST_STREAM to it so that the client closes. But the max_id is only updated on the successful path of h2c_handle_stream_new(), which may be reentered for partial frames or CONTINUATION frames, and as a result we don't increment it if an extraneous stream ID is rejected. Normally it doesn't have any consequence. But on a POST it can have some if the DATA frame immediately follows the faulty HEADERS frame: with max_id not incremented, the stream remains in IDLE state, and the DATA frame now lands in an invalid state from a protocol's perspective, which must lead to a connection error instead of a stream error. This can be tested by modifying the code to send an arbitrarily large MAX_CONCURRENT_STREAM setting and using h2load to send more concurrent streams than configured: with a GET, only a tiny fraction of them will report an error (e.g. 101 streams for 100 accepted will result in ~1% failure), but when sending data, most of the streams will be reported as failed because the connection will be closed. By updating the max_id earlier, the stream is now considered as closed when the DATA frame arrives and it's silently discarded. This must be backported to all versions but only if the code is exactly the same. Under no circumstance this ID may be updated for a partial frame (i.e. only update it before or just after calling h2c_frt_steam_new()).	2022-05-12 09:29:58 +02:00
Emeric Brun	314e6ec822	BUG/MAJOR: dns: multi-thread concurrency issue on UDP socket This patch adds a lock on the struct dgram_conn to ensure that an other thread cannot trash a fd or alter its status while the current thread processing it on for send/receive/connect operations. Starting with the 2.4 version this could cause a crash when a DNS request is failing, setting the FD of the dgram structure to -1. If the dgram structure is reused after that, a read access to fdtab[-1] is attempted. The crash was only triggered when compiled with ASAN. In previous versions the concurrency issue also exists but is less likely to crash. This patch must be backported until v2.4 and should be adapt for v < 2.4.	2022-05-11 15:20:10 +02:00
Willy Tarreau	63fc900ba2	BUG/MEDIUM: ssl: fix the gcc-12 broken fix :-( ... or how a bogus warning forces you to do tricky changes in your code and fail on a length test condition! Fortunately it changed in the right direction that immediately broke, due to a missing "> sizeof(path)" that had to be added to the already ugly condition. This fixes recent commit `393e42ae5` ("BUILD: ssl: work around bogus warning in gcc 12's -Wformat-truncation"). It may have to be backported if that one is backported.	2022-05-09 21:16:13 +02:00
Willy Tarreau	d86793479f	BUILD: listener: shut report of possible null-deref in listener_accept() When building without threads, gcc 12 says that there's a null-deref in _HA_ATOMIC_INC() called from listener_accept(). It's just that the code was originally written in an attempt not to always have a proxy for a listener and that there are two places where the pointer is tested before being used, so the compiler concludes that the pointer might be null hence that other places are null-derefs. In practice the pointer cannot be null there (and never has been), but since that code was initially built that way and it's only a matter of adding a pair of braces to shut it up, let's respect that initial attempt in case one day we need it. This one was also reported by Ilya in issue #1513, though with threads enabled in his case. This may have to be backported if users complain about new breakage with gcc-12.	2022-05-09 20:49:36 +02:00
Willy Tarreau	393e42ae5f	BUILD: ssl: work around bogus warning in gcc 12's -Wformat-truncation As was first reported by Ilya in issue #1513, Gcc 12 incorrectly reports a possible overflow from the concatenation of two strings whose size was previously checked to fit: src/ssl_crtlist.c: In function 'crtlist_parse_file': src/ssl_crtlist.c:545:58: error: '%s' directive output may be truncated writing up to 4095 bytes into a region of size between 1 and 4096 [-Werror=format-truncation=] 545 \| snprintf(path, sizeof(path), "%s/%s", global_ssl.crt_base, crt_path); \| ^~ src/ssl_crtlist.c:545:25: note: 'snprintf' output between 2 and 8192 bytes into a destination of size 4097 545 \| snprintf(path, sizeof(path), "%s/%s", global_ssl.crt_base, crt_path); \| ^~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~ It would be a bit concerning to disable -Wformat-truncation because it might detect real programming mistakes at other places. The solution adopted in this patch is absolutely ugly and error-prone, but it works, it consists in integrating the snprintf() call in the error condition and to test the result again. Let's hope a smarter compiler will not warn that this test is absurd since guaranteed by the first condition... This may have to be backported for those suffering from a compiler upgrade.	2022-05-09 20:32:11 +02:00
Remi Tricot-Le Breton	444d702130	BUG/MINOR: ssl: Fix typos in crl-file related CLI commands The CRL file CLI update code was strongly based off the CA one and some copy-paste issues were then introduced. This patch fixes GitHub issue #1685. It should be backported to 2.5.	2022-05-09 14:23:04 +02:00
William Lallemand	589570df1f	MEDIUM: ssl: ignore dotfiles when loading a dir w/ crt Ignore the files starting with a dot when trying to load a directory with the "crt" directive. Should fix issue #1689.	2022-05-09 10:41:51 +02:00
William Lallemand	e4b93eb947	MINOR: ssl: ignore dotfiles when loading a dir w/ ca-file Ignore the files starting with a dot when trying to load a directory with the "ca-file directive".	2022-05-09 09:33:25 +02:00
Willy Tarreau	6ef1648dc2	CLEANUP: stats: rename the stats state values an mark the old ones deprecated The STAT_ST_* values have been abused by virtually every applet and CLI keyword handler, and this must not continue as it's a source of bugs and of overly complicated code. This patch renames the states to STAT_STATE_*, and keeps the previous enum while marking each entry as deprecated. This should be sufficient to catch out-of-tree code that might rely on them and to let them know what to do with that.	2022-05-06 18:33:49 +02:00
Willy Tarreau	1c0715b12a	CLEANUP: cli: move the status print context into its own context Now that the CLI's print context is alone in the appctx, it's possible to refine the appctx's ctx layout so that the cli part matches exactly a regular svcctx, and as such move the CLI context into an svcctx like other applets. External code will still build and work because the struct cli perfectly maps onto the struct cli_print_ctx that's located into svc.storage. This is of course only to make a smooth transition during 2.6 and will disappear immediately after. A tiny change had to be applied to the opentracing addon which performs direct accesses to the CLI's err pointer in its own print function. The rest uses the standard cli_print_* which were the only ones that needed a small change. The whole "ctx.cli" struct could be tagged as deprecated so that any possibly existing external code that relies on it will get build warnings, and the comments in the struct are pretty clear about the way to fix it, and the lack of future of this old API.	2022-05-06 18:33:22 +02:00
Willy Tarreau	aa229ccc4c	MINOR: lua: move the http service context out of appctx.ctx Just like for the TCP service, let's move the context away from appctx.ctx. A new struct hlua_http_ctx was defined, reserved in hlua_applet_http_init() and used everywhere else. Similarly, the task dump code will no more report decoded stack traces in case these services would be involved. That may be solved later.	2022-05-06 18:13:36 +02:00
Willy Tarreau	e23f33bbfe	MINOR: lua: move the tcp service storage outside of appctx.ctx The use-service mechanism for Lua in TCP mode relies on the hlua_tcp storage in appctx->ctx. We can move its definition to hlua.c and simply use appctx_reserve_svcctx() to reserve and access the stoage. One tiny side effect is that the task dump used in panics will not show anymore the Lua call stack in its trace. For this a better API is needed from the Lua code to expose a function that does the job from an appctx.	2022-05-06 18:13:36 +02:00
Willy Tarreau	5321da9df0	MEDIUM: lua: move the cosocket storage outside of appctx.ctx The Lua cosockets were using appctx.ctx.hlua_cosocket. Let's move this to a local definition of "hlua_csk_ctx" in hlua.c, which is allocated from the appctx by hlua_socket_new(). There's a notable change which is that, while previously the xref link with the peer was established with the appctx, it's now in the hlua_csk_ctx. This one must then hold a pointer to the appctx. The code was adjusted accordingly, and now that part of the code doesn't use the appctx.ctx anymore.	2022-05-06 18:13:36 +02:00
Willy Tarreau	f61494c708	CLEANUP: cache: take the context out of appctx.ctx The context was moved to a local definition in the cache code, and there's nothing specific to the cache anymore in the appctx. The struct is stored into the appctx's storage area via the svcctx.	2022-05-06 18:13:36 +02:00
Willy Tarreau	c7afedc140	BUILD: applet: mark the appctx's st2 variable as deprecated This one has been misused for a while as well, it's time to deprecate it since we don't use it anymore. It will be removed in 2.7 and for now is only marked as deprecated. Since we need to guarantee that it's zeroed before starting any applet or CLI command, it was moved into an anonymous union where its sibling is not marked as deprecated so that we can continue to initialize it without triggering a warning. If you found this commit after a bisect session you initiated to figure why you got some build warnings and don't know what to do, have a look at the code that deals with the "show fd", "show sess" or "show servers" commands, as it's supposed to be self-explanatory about the tiny changes to apply to your code to port it. If you find APPLET_MAX_SVCCTX to be too small for your use case, either kindly ask for a tiny extension (and try to get your code merged), or just use a pool.	2022-05-06 18:13:36 +02:00
Willy Tarreau	23a2407843	CLEANUP: spoe: do not use appctx.ctx anymore The spoe code already uses its own generic pointer, let's move it to svcctx instead of keeping a struct spoe in the appctx union.	2022-05-06 18:13:36 +02:00
Willy Tarreau	455caef642	CLEANUP: peers: do not use appctx.ctx anymore The peers code already uses its own generic pointer, let's move it to svcctx instead of keeping a struct peers in the appctx union.	2022-05-06 18:13:36 +02:00
Willy Tarreau	1eea6657fb	CLEANUP: httpclient: do not use the appctx.ctx anymore The httpclient already uses its own pointer and only used to store this single pointer into the appctx.ctx field. Let's just move it to the svcctx and remove this entry from the appctx union.	2022-05-06 18:13:36 +02:00
Willy Tarreau	89a7c41e24	CLEANUP: httpclient/cli: use a locally-defined context instead of ctx.cli The httpclient's CLI uses ctx.cli.i0 for its flags and .p0 for the client instance. Let's have a locally defined structure for this so that we don't need the generic cli variables anymore.	2022-05-06 18:13:36 +02:00
Willy Tarreau	b128f49d89	CLEANUP: cli: make "show cli sockets" use its own context Let's create a show_sock_ctx to store the bind_conf and the listener. The entry is reserved when entering the I/O handler since there's no parser here. That's fine because the function doesn't touch the area.	2022-05-06 18:13:36 +02:00
Willy Tarreau	4df54eb151	CLEANUP: cli: simplify the "show cli sockets" I/O handler The code is was a bit convoluted by the use of a state machine around st2 that is not used since only the STAT_ST_LIST state was used, and the test of global.cli_fe inside the loop while it ought better be tested before entering there. Let's get rid of this unneded state and simplify the code. There's no more need for ->st2 now. The code looks more changed than it really is due to the reindent caused by the removal of the switch statement, but "git show -b" shows what really changed.	2022-05-06 18:13:36 +02:00
Willy Tarreau	307dbb33bb	CLEANUP: cli: make "show env" use its own context There is the variable to start from (or environ) and an option to stop after dumping the first one, just like "show fd". Let's have a small locally-defined context with these two fields.	2022-05-06 18:13:36 +02:00
Willy Tarreau	741a5a9cb4	CLEANUP: cli: make "show fd" use its own context The "show fd" command used to rely on cli.i0 for the fd, and st2 just to decide whether to stop after the first value or not. It could have been possible to decide to use just a negative integer to dump a single value, but it's as easy and more durable to declare a two-field struct show_fd_ctx for this.	2022-05-06 18:13:36 +02:00
Willy Tarreau	acf6a44908	CLEANUP: proxy/cli: make "show backend" only use the generic context Let's use appctx->svcctx instead of abusing cli.p0 to store the current proxy being dumped.	2022-05-06 18:13:36 +02:00
Willy Tarreau	d741e9c4b1	CLEANUP: proxy/cli: get rid of appctx->st2 in "show servers" Now that we have the show_srv_ctx, let's store a state in it. We only need two states here, header and list.	2022-05-06 18:13:36 +02:00
Willy Tarreau	72a04238d5	CLEANUP: proxy/cli: make use of a locally defined context for "show servers" The command uses a pointer to the current proxy being dumped, one to the current server being dumped, an optional ID of the only proxy to dump (which is in fact used as a boolean), and a flag indicating if we're doing a "show servers conn" or a "show servers state". Let's move all this to a struct show_srv_ctx.	2022-05-06 18:13:36 +02:00
Willy Tarreau	c6dfef7a0b	CLEANUP: cache/cli: make use of a locally defined context for "show cache" The command uses a pointer to a cache instance and the next key to dump, they were in cli.p0/i0 respectively, let's move them to a struct show_cache_ctx.	2022-05-06 18:13:36 +02:00
Willy Tarreau	12d5228a44	CLEANUP: resolvers/cli: remove the unneeded appctx->st2 from "show resolvers" The command uses this state but _INIT immediately turns to _LIST, which turns to _FIN at the end without doing anything in that state, thus the only existing state is _LIST so we don't need to store a state. Let's just get rid of it.	2022-05-06 18:13:36 +02:00
Willy Tarreau	db933d6fdd	CLEANUP: resolvers/cli: make "show resolvers" use a locally-defined context The command was using cli.p0/p1/p2 to select which section to dump, the current section and the current ns. Let's instead have a locally defined "show_resolvers_ctx" section for this.	2022-05-06 18:13:36 +02:00
Willy Tarreau	6e3fc483f7	CLEANUP: ring/cli: use a locally-defined context instead of using ctx.cli The ring code was using ctx.cli.i0/p0/o0 to store its context during CLI dumps via "show events" or "show errors". Let's use a locally defined context and drop that.	2022-05-06 18:13:36 +02:00
Willy Tarreau	cba8838e59	CLEANUP: ring: pass the ring watch flags to ring_attach_cli(), not in ctx.cli The ring watch flags (wait, seek end) were dangerously passed via ctx.cli.i0 from "show buf" in sink.c:cli_parse_show_events(), or implicitly reset in "show errors". That's very unconvenient, difficult to follow, and prone to short-term breakage. Let's pass an extra argument to ring_attach_cli() to take these flags, now defined in ring-t.h as RING_WF_*, and let the function set them itself where appropriate (still ctx.cli.i0 for now).	2022-05-06 18:13:36 +02:00
Willy Tarreau	40e952f1a6	CLEANUP: debug/cli: make "debug dev memstats" not use ctx.cli anymore There was only the need for a start and a stop pointer, and a show_all flag. All of that moved to a locally-defined struct dev_mem_ctx.	2022-05-06 18:13:36 +02:00
Willy Tarreau	e06bbf3f19	CLEANUP: debug/cli: make "debug dev fd" not use ctx.cli anymore The command only requires to store an int, but it will be useful later to have a struct to pass extra info such as an "all" flag to dump all FDs. The new context is now a struct dev_fd_ctx stored in svcctx.	2022-05-06 18:13:36 +02:00
Willy Tarreau	e8d006a79a	CLEANUP: activity/cli: make "show profiling" not use ctx.cli anymore The I/O handler was using ctx.cli.i0/i1/o0/o1. Let's put all that into a locally-defined context and use it instead.	2022-05-06 18:13:36 +02:00
Willy Tarreau	42cc831abf	CLEANUP: sink: use the generic context to store the forwarder's context Instead of having a struct that contains a single pointer in the appctx context, let's directly use the generic context pointer and get rid of the now unused sft.ptr entry.	2022-05-06 18:13:36 +02:00
Willy Tarreau	0d626a5610	CLEANUP: dns: stop abusing the sink forwarder's context The DNS code was abusing the sink forwarder's context as its own. Let's make it directly use the generic context pointer instead.	2022-05-06 18:13:36 +02:00
Willy Tarreau	fa11df5d03	CLEANUP: ssl/cli: make "add ssl crtlist" not use st2 anymore Several steps are used during the addition of a crtlist to yield during long operations, and states are used for this. Let's just not use the st2 anymore and place the state inside the add_crtlist_ctx struct instead.	2022-05-06 18:13:36 +02:00
Willy Tarreau	6b6c363a6b	CLEANUP: ssl/cli: make "add ssl crtlist" use its own context This command was using cli.p0/p1/p2 in the io_handler. Let's move them to a command-specific "struct add_crtlist_ctx".	2022-05-06 18:13:36 +02:00
Willy Tarreau	a2fcca0939	CLEANUP: ssl/cli: make "{show\|dump} ssl crtlist" use its own context These commands were using cli.i0/p0/p1 and in a not very clean way since they use the same parser but with different types depending on the I/O handler. Given there was no explanation about what the variables were supposed to be, they were named based on best guess and placed into a new "show_crtlist_ctx" structure.	2022-05-06 18:13:36 +02:00
Willy Tarreau	170b35bb95	CLEANUP: ssl/cli: make "show ssl ocsp-response" not use cli.p0 anymore Instead the single-pointer context is placed into appctx->svcctx. There's no need to declare a structure there for this.	2022-05-06 18:13:36 +02:00
Willy Tarreau	9c5a38c1b8	CLEANUP: ssl/cli: make "show tlskeys" not use appctx->st2 anymore A new "state" enum was added to "show_keys_ctx" for this, and only 3 states are needed.	2022-05-06 18:13:36 +02:00
Willy Tarreau	bd33864373	CLEANUP: ssl/cli: add a new "dump_entries" field to "show_keys_ref" This gets rid of a ugly hack consisting in checking the IO handler's address while one is defined as an inline function calling the second.	2022-05-06 18:13:36 +02:00
Willy Tarreau	a938052113	CLEANUP: ssl/cli: stop using ctx.cli.i0/i1/p0 for "show tls-keys" This creates a local context of type show_keys_ctx which contains the equivalent fields with more natural names.	2022-05-06 18:13:36 +02:00
Willy Tarreau	1d6dd80d05	CLEANUP: ssl/cli: stop using appctx->st2 for "commit ssl ca/crl" A new entry "state" was added into the commit_cacrl_ctx struct instead.	2022-05-06 18:13:36 +02:00
Willy Tarreau	dec23dc43f	CLEANUP: ssl/cli: use a local context for "commit ssl {ca\|crl}file" These two commands use distinct parse/release functions but a common iohandler, thus they need to keep the same context. It was created under the name "commit_cacrlfile_ctx" and holds a large part of the pointers (6) and the ca_type field that helps distinguish between the two commands for the I/O handler. It looks like some of these fields could have been merged since apparently the CA part only uses cafile and the CRL part crlfile, while both old and new are of type cafile_entry and set only for each type. This could probably even simplify some parts of the code that tries to use the correct field. These fields were the last ones to be migrated thus the appctx's ssl context could finally be removed.	2022-05-06 18:13:36 +02:00
Willy Tarreau	a06b9a5ccf	CLEANUP: ssl/cli: use a local context for "set ssl crlfile" Just like for "set ssl cafile", the command doesn't really need this context which doesn't outlive the parsing function but it was there for a purpose so it's maintained. Only 3 fields were used from the appctx's ssl context: old_crlfile_entry, new_crlfile_entry, and path. These ones were reinstantiated into a new "set_crlfile_ctx" struct. It could have been merged with the one used in "set cafile" if the fields had been renamed since cafile and crlfile are of the same type (probably one of them ought to be renamed?). None of these fields could be dropped as they are still shared with other commands.	2022-05-06 18:13:36 +02:00
Willy Tarreau	a37693f7d8	CLEANUP: ssl/cli: use a local context for "set ssl cafile" Just like for "set ssl cert", the command doesn't really need this context which doesn't outlive the parsing function but it was there for a purpose so it's maintained. Only 3 fields were used from the appctx's ssl context: old_cafile_entry, new_cafile_entry, and path. These ones were reinstantiated into a new "set_cafile_ctx" struct. None of them could be dropped as they are still shared with other commands.	2022-05-06 18:13:36 +02:00
Willy Tarreau	329f4b4f2f	CLEANUP: ssl/cli: use a local context for "set ssl cert" The command doesn't really need any storage since there's only a parser, but since it used this context, there might have been plans for extension, so better continue with a persistent one. Only old_ckchs, new_ckchs, and path were being used from the appctx's ssl context. There ones moved to the local definition, and the two former ones were removed from the appctx since not used anymore.	2022-05-06 18:13:36 +02:00
Willy Tarreau	cb1b4ed7b6	CLEANUP: ssl/cli: stop using appctx->st2 for "commit ssl cert" A new entry "state" was added into the commit_cert_ctx struct instead.	2022-05-06 18:13:36 +02:00
Willy Tarreau	a645b6a21f	CLEANUP: ssl/cli: use a local context for "commit ssl cert" This command only really uses old_ckchs, new_ckchs and next_ckchi from the appctx's ssl context. The new structure "commit_cert_ctx" only has these 3 fields, though none could be removed from the shared ssl context since they're still used by other commands.	2022-05-06 18:13:36 +02:00
Willy Tarreau	96c9a6c752	CLEANUP: ssl/cli: use a local context for "show ssl cert" This command only really uses old_ckchs, cur_ckchs and the index in which the transaction was stored. The new structure "show_cert_ctx" only has these 3 fields, and the now unused "cur_ckchs" and "index" could be removed from the shared ssl context.	2022-05-06 18:13:36 +02:00
Willy Tarreau	f3e8b3e877	CLEANUP: ssl/cli: use a local context for "show crlfile" Now this command doesn't share any context anymore with "show cafile" nor with the other commands. The previous "cur_cafile_entry" field from the applet's ssl context was removed as not used anymore. Everything was moved to show_crlfile_ctx which only has 3 fields.	2022-05-06 18:13:36 +02:00
Willy Tarreau	50c2f1e0cd	CLEANUP: ssl/cli: use a local context for "show cafile" Saying that the layout and usage of the various variables in the ssl applet context is a mess would be an understatement. It's very hard to know what command uses what fields, even after having moved away from the mix of cli and ssl. Let's extract the parts used by "show cafile" into their own structure. Only the "show_all" field would be removed from the ssl ctx, the other fields are still shared with other commands.	2022-05-06 18:13:35 +02:00
Willy Tarreau	bcda5f6bcd	CLEANUP: hlua/cli: take the hlua_cli context definition out of the appctx This context is used by CLI keywords registered by Lua. We can take it out of the appctx and use the generic command context allocation so that the appctx doesn't have to declare a specific one anymore. The context is created during parsing.	2022-05-06 18:13:35 +02:00
Willy Tarreau	41f885241e	CLEANUP: stats/cli: stop using appctx->st2 Instead, let's have the state as an enum inside the context. It's much cleaner and safer as we know nobody else touches it.	2022-05-06 18:13:35 +02:00
Willy Tarreau	91cefcaba4	CLEANUP: stats/cli: take the "show stat" context definition out of the appctx This makes use of the generic command context allocation so that the appctx doesn't have to declare a specific one anymore. The context is created during parsing (both in the CLI and HTTP). The change looks large but it's particularly mechanical. The context initialization appears in stats.c and http_ana.c. The context is used in stats.c and resolvers.c since "show stat resolvers" points there. That's the reason why the definition moved to stats.h. "show info" and "show stat" continue to share the same state definition for now. Nothing else was modified.	2022-05-06 18:13:35 +02:00
Willy Tarreau	7bf20caacc	CLEANUP: cli: initialize the whole appctx->ctx, not just the stats part Historically the CLI was a second access to the stats and we've continued to initialize only the stats part when initializing the CLI. Let's make sure we do that on the whole ctx instead. It's probably not more needed at all nowadays but better stay on the safe side.	2022-05-06 18:13:35 +02:00
Willy Tarreau	ce9123c005	CLEANUP: peers/cli: remove unneeded state STATE_INIT All the settings in this initial state are konwn at parsing time, there's no need for an initial state to bootstrap other ones.	2022-05-06 18:13:35 +02:00
Willy Tarreau	3a31e37518	CLEANUP: peers/cli: stop using appctx->st2 for the dump state Let's instead define a 4-state enum solely for this use case, and place it into the command's context. Note that END and FIN were already aliases, which is why they were merged.	2022-05-06 18:13:35 +02:00
Willy Tarreau	cb8bf17900	CLEANUP: peers/cli: take the "show peers" context definition out of the appctx This makes use of the generic command context allocation so that the appctx doesn't have to declare a specific one anymore. The context is created during parsing. The code also uses st2 which deserves being addressed in separate commit.	2022-05-06 18:13:35 +02:00
Willy Tarreau	c7e9706e0f	CLEANUP: map/cli: always detach the backref from the list after "show map" There's no point checking the state before deciding to detach the backref on "show map", it should always be done if the list is not empty. Note that being empty guarantees that it's not linked into the list, and conversely not being empty guarantees that it's in the list, hence the test doesn't need to be performed under the lock.	2022-05-06 18:13:35 +02:00
Willy Tarreau	a0d6280af4	CLEANUP: map/cli: stop using appctx->st2 for the dump state Let's instead define a 3-state enum solely for this use case, and place it into the command's context.	2022-05-06 18:13:35 +02:00
Willy Tarreau	76f771e78e	CLEANUP: map/cli: stop using cli.i0/i1 to store the generation numbers The "show map"/"clear map" code used to rely on the cli's i0/i1 fields to store the generation numbers to work with. That's particularly dirty because it's done at places where ctx.map is also manipulated while they are part of the same union, and the reason why this didn't cause trouble is because cli.i0/i1 are at offset 216/224 while the map parts end at 204, so luckily there was no overlap. Let's add these fields to the map context.	2022-05-06 18:13:35 +02:00
Willy Tarreau	0fcecc63c8	CLEANUP: map/cli: take the "show map" context definition out of the appctx This makes use of the generic command context allocation so that the appctx doesn't have to declare a specific one anymore. The context is created during parsing. Many commands, including pure parsers, use this context but that's not a problem as it's designed to be used this way. Due to this, many lines are changed but that's in fact a replacement of "appctx->ctx.map" with "ctx->". Note that the code also uses st2 which deserves being addressed in separate commit.	2022-05-06 18:13:35 +02:00
Willy Tarreau	0f154ed1ef	CLEANUP: stick-table/cli: remove the unneeded STATE_INIT for "show table" This state is pointless now, it just serves to initialize the initial table pointer while this can be done more easily in the parser, so let's do that and drop that state.	2022-05-06 18:13:35 +02:00
Willy Tarreau	7849d4c912	CLEANUP: stick-table/cli: stop using appctx->st2 for the dump state Let's instead define a 4-state enum solely for this use case, and place it into the command's context. Note that END and FIN were already aliases, which is why they were merged.	2022-05-06 18:13:35 +02:00
Willy Tarreau	3c69e08e96	CLEANUP: stick-table/cli: take the "show table" context definition out of the appctx This makes use of the generic command context allocation so that the appctx doesn't have to declare a specific one anymore. The context is created during parsing. The code also uses st2 which deserves being addressed in separate commit.	2022-05-06 18:13:35 +02:00
Willy Tarreau	0fd8f0e236	CLEANUP: proxy/cli: take the "show errors" context definition out of the appctx This makes use of the generic command context allocation so that the appctx doesn't have to declare a specific one anymore. The context is created during parsing. The code still has room for improvement, such as in the "flags" field where bits are hard-coded, but they weren't modified.	2022-05-06 18:13:35 +02:00
Willy Tarreau	6177cfc3b5	CLEANUP: stream/cli: remove the now unneeded dump state from "show sess" The state was a constant, let's remove what remains of the switch/case. The code from the "case" statement was only reindented as can be checked with "git show -b".	2022-05-06 18:13:35 +02:00
Willy Tarreau	bb4e289fa9	CLEANUP: stream/cli: remove the unneeded STATE_FIN state from "show sess" This state is only an alias for "thr >= global.nbthread" which is the sole condition that indicates the end and reaches it. Let's just drop this state. There's only the STATE_LIST that's left.	2022-05-06 18:13:35 +02:00
Willy Tarreau	f3629f88ac	CLEANUP: stream/cli: remove the unneeded init state from "show sess" This state was only used to preset the list element. Now that we can guarantee that the context can be properly preset during the parsing we don't need this state anymore. The first pointer has to be set to point to the first stream during the initial call which is detected by the pointer not yet being set (null). Thanks to this we can also remove one state check on the abort path.	2022-05-06 18:13:35 +02:00
Willy Tarreau	7fb591a4e4	CLEANUP: stream/cli: stop using appctx->st2 for the dump state Let's instead define a 3-state enum solely for this use case, and place it into the command's context.	2022-05-06 18:13:35 +02:00
Willy Tarreau	39f097d965	CLEANUP: stream/cli: take the "show sess" context definition out of the appctx This makes use of the generic command context allocation so that the appctx doesn't have to declare a specific one anymore. The context is created during parsing.	2022-05-06 18:13:35 +02:00
Willy Tarreau	009e42bc59	CLEANUP: applet: make appctx_new() initialize the whole appctx Till now, appctx_new() used to allocate an entry from the pool and pass it through appctx_init() to initialize a debatable part of it that didn't correspond anymore to the comments, and fill other fields. It's hard to say what is fully initialized and what is not. Let's get rid of that, and always zero the initialization (appctx are not that big anyway, even with the cache there's no difference in performance), and initialize what remains. this is cleaner and more resistant to new field additions. The appctx_init() function was removed.	2022-05-06 18:13:35 +02:00
Willy Tarreau	f12f32a0fa	MINOR: applet: reserve some generic storage in the applet's context Instead of using existing fields and having to put keyword-specific contexts in the applet definition, let's have the appctx provide a generic storage area that's currently large enough for existing CLI commands and small applets, and a function to allocate that storage. The function will be responsible for verifying that the requested size fits in the area so that the caller doesn't need to add specific checks; it is validated during development as this size is static and will not change at runtime. In addition the caller doesn't even need to free() the area since it's part of an existing context. For the caller's convenience, a context pointer "svcctx" for the command is also provided so that the allocated area can be placed there (or possibly any other one in case a larger area is needed). The struct's layout has been temporarily complicated by adding one level of anonymous union on top of the "ctx" one. This will allow us to preserve "ctx" during 2.6 for compatibility with possible external code and get rid of it in 2.7. This explains why the diff extends to the whole "ctx" union, but a "git show -b" shows that only one extra layer was added. In order to make both the svcctx pointer and its storage accessible without further enlarging the appctx structure, both svcctx and the storage share the same storage as the ctx part. This is done by having them placed in the union with a protected overlapping area for svcctx, for which a shadow member is also present in the storage area: union { void* svcctx; // variable accessed by services struct { void shadow; // shadow of svcctx; char storage[]; // where most services store their data }; union { // older commands store here and ignore svcctx ... } ctx; }; I.e. new applications will use appctx->svcctx while older ones will be able to continue to use appctx->ctx. The whole area (including the pointer's context) is zeroed before any applet is initialized, and before CLI keyword processor's first invocation, as it is an important part of the existing keyword processors, which makes CLI keywords effectively behave like applets.	2022-05-06 18:13:35 +02:00
Willy Tarreau	1b948ef426	CLEANUP: ssl/cli: do not loop on unknown states in "add ssl crt-list" handler The io_handler in "add ssl crt_list" is built around a "while" loop that only makes forward progress and that doesn't handle its final state as it's not supposed to be called again once reached. This makes the code confusing because its construct implies an infinite loop for such a state (or any other unhandled one). Let's just remove that unneeded loop.	2022-05-06 18:13:35 +02:00
Willy Tarreau	4fd9b4ddf0	BUG/MINOR: ssl/cli: fix "show ssl cert" not to mix cli+ssl contexts The "show ssl cert" command mixes some generic pointers from the "ctx.cli" struct with context-specific ones from "ctx.ssl" while both are in a union. Amazingly, despite the use of both p0 and i0 to store respectively a pointer to the current ckchs and a transaction id, there was no overlap with the other pointers used during these operations, but should these fields be reordered or slightly updated this will break. Comments were added above the faulty functions to indicate which fields they are using. This needs to be backported to 2.5.	2022-05-06 18:13:35 +02:00
Willy Tarreau	4cf3ef8007	BUG/MINOR: ssl/cli: fix "show ssl crl-file" not to mix cli+ssl contexts The "show ssl crl-file" command mixes some generic pointers from the "ctx.cli" struct with context-specific ones from "ctx.ssl" while both are in a union. It's fortunate that the p1 pointer in use is located before the first one used (it overlaps with old_cafile_entry). But should these fields be reordered or slightly updated this will break. This needs to be backported to 2.5.	2022-05-06 18:13:35 +02:00
Willy Tarreau	06305798f7	BUG/MINOR: ssl/cli: fix "show ssl ca-file <name>" not to mix cli+ssl contexts The "show ssl ca-file <name>" command mixes some generic pointers from the "ctx.cli" struct and context-specific ones from "ctx.ssl" while both are in a union. The i0 integer used to store the current ca_index overlaps with new_crlfile_entry which is thus harmless for now but is at the mercy of any reordering or addition of these fields. Let's add dedicated fields into the ssl structure for this. Comments were added on top of the affected functions to indicate what they use. This needs to be backported to 2.5.	2022-05-06 18:13:35 +02:00
Willy Tarreau	821c3b0b5e	BUG/MINOR: ssl/cli: fix "show ssl ca-file/crl-file" not to mix cli+ssl contexts The "show ca-file" and "show crl-file" commands mix some generic pointers from the "ctx.cli" struct and context-specific ones from "ctx.ssl" while both are in a union. It's fortunate that the p0 pointer in use is located immediately before the first one used (it overlaps with next_ckchi_link, and old_cafile_entry is safe). But should these fields be reordered or slightly updated this will break. Comments were added on top of the affected functions to indicate what they use. This needs to be backported to 2.5.	2022-05-06 18:13:35 +02:00
Willy Tarreau	1ae0c43244	BUG/MINOR: map/cli: make sure patterns don't vanish under "show map"'s init When "show map" initializes itself, it first takes the reference to the starting point under a lock, then releases it before switching to state STATE_LIST, and takes the lock again. The problem is that it is possible for another thread to remove the first element during this unlock/lock sequence, and make the list run anywhere. This is of course extremely unlikely but not impossible. Let's initialize the pointer in the STATE_LIST part under the same lock, which is simpler and more reliable. This should be backported to all versions.	2022-05-06 18:13:35 +02:00
Willy Tarreau	2edaace575	BUG/MINOR: map/cli: protect the backref list during "show map" errors In case of write error in "show map", the backref is detached but the list wasn't locked when this is done. The risk is very low but it may happen that two concurrent "show map" one of which would fail or one "show map" failing while the same entry is being updated could cause a crash. This should be backported to all stable versions.	2022-05-06 18:13:35 +02:00
Willy Tarreau	4f9f157537	BUG/MINOR: proxy/cli: don't enumerate internal proxies on "show backend" Commit `e7f74623e` ("MINOR: stats: don't output internal proxies (PR_CAP_INT)") in 2.5 ensured we don't dump internal proxies on the stats page, but the same is needed for "show backend", as since the addition of the HTTP client it now appears there: $ socat /tmp/sock1 - <<< "show backend" # name <HTTPCLIENT> This needs to be backported to 2.5.	2022-05-06 18:13:35 +02:00
Willy Tarreau	241a006d79	BUG/MEDIUM: cli: make "show cli sockets" really yield This command was introduced in 1.8 with commit `eceddf722` ("MEDIUM: cli: 'show cli sockets' list the CLI sockets") but its yielding doesn't work. Each time it enters, it restarts from the last bind_conf but enumerates all listening sockets again, thus it loops forever. The risk that it happens in field is low but it easily triggers on port ranges after 400-500 sockets depending on the length of their addresses: global stats socket /tmp/sock1 level admin stats socket 192.168.8.176:30000-31000 level operator $ socat /tmp/sock1 - <<< "show cli sockets" (...) ipv4@192.168.8.176:30426 operator all ipv4@192.168.8.176:30427 operator all ipv4@192.168.8.176:30428 operator all ipv4@192.168.8.176:30000 operator all ipv4@192.168.8.176:30001 operator all ipv4@192.168.8.176:30002 operator all ^C This patch adds the minimally needed restart point for the listener so that it can easily be backported. Some more cleanup is needed though.	2022-05-06 18:13:35 +02:00
Willy Tarreau	4e047e7d0e	BUG/MEDIUM: resolvers: make "show resolvers" properly yield The "show resolvers" command is bogus, it tries to implement a yielding mechanism except that if it yields it restarts from the beginning, until it manages to fill the buffer with only line breaks, and faces error -2 that lets it reach the final state and exit. The risk is low since it requires about 50 name servers to reach that state, but it's not impossible, especially when using multiple sections. In addition, the extraneous line breaks, if sent over an interactive connection, will desynchronize the commands and make the client believe the end was reached after the first nameserver. This cannot be fixed separately because that would turn this bug into an infinite loop since it's the line feed that manages to fill the buffer and stop it. The fix consists in saving the current resolvers section into ctx.cli.p1 and the current nameserver into ctx.cli.p2. This should be backported, but that code moved a lot since it was introduced and has always been bogus. It looks like it has mostly stabilized in 2.4 with commit `c943799c86` so the fix might be backportable to 2.4 without too much effort.	2022-05-06 18:13:35 +02:00
William Lallemand	89e236f246	BUG/MINOR: startup: usage() when no -cc arguments Exit correctly with usage() instead of segfaulting when no argument were passed to -cc. Must be backported in 2.5.	2022-05-06 17:22:36 +02:00
William Lallemand	7867f63313	MEDIUM: resolvers: create a "default" resolvers section at startup Try to create a "default" resolvers section at startup, but does not display any error nor warning. This section is initialized using the /etc/resolv.conf of the system. This is opportunistic and with no guarantee that it will work (but it should on most systems). This is useful for the httpclient as it allows to use the DNS resolver without any configuration in most of the cases. The function is called from the httpclient_pre_check() function to ensure than we tried to create the section before trying to initiate the httpclient. But it is also called from the resolvers.c to ensure the section is created when the httpclient init was disabled.	2022-05-06 17:02:15 +02:00
William Lallemand	0164a40c59	BUG/MINOR: tcp/http: release the expr of set-{src,dst}[-port] Release the expression used by the set-{src,dst}[-port] actions so we keep valgrind happy upon an exit or an haproxy -c. Could be backported in every supported version.	2022-05-06 17:02:15 +02:00
Willy Tarreau	7831e0272e	BUILD: debug: unify the definition of ha_backtrace_to_stderr() It was both defined as ha_backtrace_to_stderr(void) and ha_backtrace_to_stderr(), and tcc is not happy with this, so let's adjust this tiny detail.	2022-05-06 15:16:19 +02:00
William Lallemand	e7f5776800	MINOR: resolvers: resolvers_new() create a resolvers with default values Split the creation of the resolve structure from the parser to resolvers_new();	2022-05-05 18:27:48 +02:00
William Lallemand	73edfe402e	MINOR: resolvers: move the resolv.conf parser in parse_resolv_conf() Move the resolv.conf parser from the cfg_parse_resolvers so it could be used separately. Some changes were made in the memprintf in order to use a char ** instead of a char *. Also the variable is tested before each memprintf so could skip them if no warnmsg nor errmsg were set.	2022-05-05 17:38:48 +02:00
William Lallemand	106bd29dd0	MINOR: resolvers: cleanup alert/warning in parse-resolve-conf Cleanup the alert and warning handling in the "parse-resolve-conf" parser to use the errmsg and warnmsg variables and memprintf. This will allow to split the parser and shut the alert/warning if needed.	2022-05-05 17:33:42 +02:00
Christopher Faulet	d934e8d963	BUG/MEDIUM: mux-h1: Be able to handle trailers when C-L header was specified The commit `2eb5243e7` ("BUG/MEDIUM: mux-h1: Set outgoing message to DONE when payload length is reached") introduced a regression. An internal error is reported when we try to forward a message with trailers while the content-length header was specified. Indeed, this case does not exist for H1 messages but it is possible in H2. This patch should solve the issue #1684. It must be backported as far as 2.4.	2022-05-05 09:39:43 +02:00
Christopher Faulet	2db904e86c	BUG/MEDIUM: mux-fcgi: Be sure to never set EOM flag on an empty HTX message This bug was already fixed at many places (stats, promex, lua) but the FCGI multiplexer is also affected. When there is no content-length specified in the response and when the END_REQUEST record is delayed, the response may be truncated because an abort is erroneously detected. If the connection is not closed because "keep-conn" option is set, the response is aborted at the end of the server timeout. This bug is a design issue with the HTX. It should be addressed. But it will probably not be possible to backport them as far as 2.4. So, for now, the only solution is to explicitly add an EOT block with the EOM flag in this case. This patch should fix the issue #1682. It must be backported as far as 2.4.	2022-05-05 09:24:55 +02:00
Christopher Faulet	c41f93c5cd	BUG/MEDIUM: conn-stream: Only keep app layer flags of the endpoint on reset The commit `a6c4a4834` ("BUG/MEDIUM: conn-stream: Don't erase endpoint flags on reset") was too laxy on reset. Only app layer flags must be preserved. On reset, the endpoint is detached. Thus all flags set by the endpoint itself or concerning its type must be removed. Without this fix, we can experienced crashes when a stream is released while a server connection attempt failed. Indeed, in this case, endpoint of the backend conn-stream is reset. But the endpoint type is still set. Thus when the stream is released, the endpoint is detached again. This patch is 2.6-specific. No backport needed. This commit depends on the previous one ("MINOR: conn-stream: Add mask from flags set by endpoint or app layer").	2022-05-05 09:23:44 +02:00
William Lallemand	7c5a7ef32b	MINOR: httpclient: allow ipv4 or ipv6 preference for resolving The httpclient.resolvers.prefer global keyword allows to configure an ipv4 or ipv6 preference when resolving. This could be useful in environment where the ipv6 is not supported.	2022-05-04 16:14:42 +02:00
William Lallemand	8a734cbae3	MINOR: httpclient: configure the resolvers section to use By default the httpclient uses the resolvers section whose ID is "default", the httpclient.resolvers.id global option allows to configure another section to use.	2022-05-04 16:14:35 +02:00
William Lallemand	683fbb86be	MINOR: httpclient: allow to configure the ca-file The global keyword httpclient.ssl.ca-file allows to configure the ca-file used for the httpclient verify.	2022-05-04 16:14:35 +02:00
William Lallemand	6fce46a910	MEDIUM: httpclient: hard-error when SSL is configured The hard_error_ssl flag is set when the configuration is explicitely done for the ssl in the httpclient. If no configuration was made, the features are simply disabled and no alert is emitted.	2022-05-04 16:13:17 +02:00
William Lallemand	85af49c5c8	MINOR: httpclient: cleanup the error handling in init Cleanup the error handling in the initialization so we rely on the ERR_CODE and use memprintf() to set the errmsg before printing it at the end of the functions.	2022-05-04 14:33:57 +02:00
William Lallemand	8b9a2df969	MINOR: init: exit() after pre-check upon error Add a test on the err_code variable so we don't go further if one of the pre-check callback failed.	2022-05-04 14:29:46 +02:00
William Lallemand	9ff95e2269	MINOR: httpclient: rename dash by dot in global option Rename the httpclient-ssl-verify into httpclient.ssl.verify.	2022-05-04 13:52:29 +02:00
William Lallemand	853327390e	MINOR: httpclient: handle unix and other socket types in dst httpclient_set_dst() allows one to set the destination address instead of using the one in the URL or resolving one from the host. This function also support other types of socket like sockpair@, unix@, anything that could be used on a server line. In order to still support this behavior, the address must be set on the backend in this particular case because the frontend connection does not support anything other than ipv4 or ipv6.	2022-05-04 11:21:01 +02:00
William Lallemand	0e23526a2c	CLEANUP: httpclient: remove the comment about resolving remove the comment of httpclient_start which states there is no resolver.	2022-05-04 11:21:01 +02:00
William Lallemand	1218d19921	MEDIUM: httpclient: allow address and port change for resolving To allow the http-request set-dst to work for the httpclient DNS resolving, some changes need to be done: - The destination address need to be set in the frontend (s->csf->dst) instead of the backend (s->csb->dst) to be able to use tcp_action_req_set_dst() - SRV_F_MAPPORTS need to be set on the proxy in order to allow the port change in alloc_dst_address()	2022-05-04 11:21:01 +02:00
William Lallemand	5392ff6e3c	MEDIUM: httpclient: http-request rules for resolving The httpclient_resolve_init() adds http-request actions which does the resolving using the Host header of the HTTP request. The parse_http_req_cond function is directly used over an array of http rules. The do-resolve rule uses the "default" resolvers section. If this section does not exist in the configuration, the resolving is disabling.	2022-05-04 11:21:01 +02:00
William Lallemand	7f1df8ff1a	MEDIUM: httpclient: remove url2sa to use a more flexible parser The httpclient DNS resolver will need a more efficient URL parser which splits the URL into parts and does not try to resolve. httpclient_spliturl uses the http_uri_parser in order to split the URL into several ist. The result of the function host part is then processed into str2ip2(), and fails if it it not an IP, allowing us to resolve the domain later.	2022-05-04 11:21:01 +02:00
Frédéric Lécaille	b57de07a21	BUG/MINOR: mux_quic: Dropped packet upon retransmission for closed streams We rely on the largest ID which was used to open streams to know if the stream we received STREAM frames for is closed or not. If closed, we return the same status as the one for a STREAM frame which was a already received one for on open stream.	2022-05-03 10:13:40 +02:00
Frédéric Lécaille	d62240c9e5	BUG/MINOR: quic: Dropped retransmitted STREAM frames It is possible that we continue to receive retransmitted STREAM frames after the mux have been released. We rely on the ->rx.streams[].nb_streams counter to check the stream was closed. If not, at this time we drop the packet.	2022-05-03 10:13:40 +02:00
Frédéric Lécaille	664741e1c5	MINOR: quic: Make the quic_conn be aware of the number of streams This is required when the retransmitted frame types when the mux is released. We add a counter for the number of streams which were opened or closed by the mux. After the mux has been released, we can rely on this counter to know if the STREAM frames are retransmitted ones or not.	2022-05-03 10:13:40 +02:00
Willy Tarreau	0367b4cf63	MINOR: session: get rid of the now unused SESS_FL_ADDR_*_SET flags That's similar to what was done for conn_streams and connections. The flags were only set exactly when the relevant pointers were allocated, so better test the pointer than the flag and stop setting the flag.	2022-05-02 17:51:51 +02:00
Willy Tarreau	030b3e6bcc	MINOR: connection: get rid of the CO_FL_ADDR_*_SET flags Just like for the conn_stream, now that these addresses are dynamically allocated, there is no single case where the pointer is set without the corresponding flag, and the flag is used as a permission to dereference the pointer. Let's just replace the test of the flag with a test of the pointer and remove all flag assignment. This makes the code clearer (especially in "if" conditions) and saves the need for future code to think about properly setting the flag after setting the pointer.	2022-05-02 17:47:46 +02:00
Willy Tarreau	158b6cf102	CLEANUP: protocol: make sure the connect_* functions always receive a dst Some of the protocol-level ->connect() functions currently dereference the connection's destination address while others test it and return an error. There's normally no more non-bogus code path that calls such functions without a valid destination address on the connection, so let's unify these functions and just place a BUG_ON() there, and drop the useless test that's supposed to return an internal error.	2022-05-02 17:47:31 +02:00
Willy Tarreau	03bd3952a6	MEDIUM: stream: remove the confusing SF_ADDR_SET flag This flag is no longer needed now that it must always match the presence of a destination address on the backend conn_stream. Worse, before previous patch, if it were to be accidently removed while the address is present, it could result in a leak of that address since alloc_dst_address() would first be called to flush it. Its usage has a long history where addresses were stored in an area shared with the connection, but as this is no longer the case, there's no reason for putting this burden onto application-level code that should not focus on setting obscure flags. The only place where that made a small difference is in the dequeuing code in case of queue redistribution, because previously the code would first clear the flag, and only later when trying to deal with the queue, would release the address. It's not even certain whether there would exist a code path going to connect_server() without calling pendconn_dequeue() first (e.g. retries on queue timeout maybe?). Now the pendconn_dequeue() code will rely on SF_ASSIGNED to decide to clear and release the address, since that flag is always set while in a server's queue, and its clearance implies that we don't want to keep the address. At least it remains consistent and there's no more risk of leaking it.	2022-05-02 16:56:01 +02:00
Willy Tarreau	b3f0d42a1d	CLEANUP: backend: make alloc_{bind,dst}_address() idempotent These functions dynamically allocate a source or destination address but start by clearing the previous one. There's a non-null risk of leaking addresses there in case of misuse. Better have them do nothing if the address was already allocated.	2022-05-02 16:20:36 +02:00
Amaury Denoyelle	291ee25696	BUG/MINOR: h3: fix parsing of unknown frame type with null length HTTP/3 implementation must ignore unknown frame type to support protocol evolution. Clients can deliberately use unknown type to test that the server is conformant : this principle is called greasing. Quiche client uses greasing on H3 frame type with a zero length frame. This reveals a bug in H3 parsing code which causes the transfer to be interrupted. Fix this by removing the break statement on ret variable. Now the parsing loop is only interrupted if input buffer is empty or the demux is blocked. This should fix http/3 freeze transfers with the quiche client. Thanks to Lucas Pardue from Cloudflare for his report on the bug. Frédéric Lecaille quickly found the source of the problem which helps me to write this patch.	2022-05-02 11:36:42 +02:00
Amaury Denoyelle	f1fc0b393b	MINOR: mux-quic: support full request channel buffer If the request channel buffer is full, H3 demuxing must be interrupted on the stream until some read is performed. This condition is reported if the HTX stream buffer qcs.rx.app_buf is full. In this case, qcs instance is marked with a new flag QC_SF_DEM_FULL. This flag cause the H3 demuxing to be interrupted. It is cleared when the HTX buffer is read by the conn-stream layer through rcv_buf operation. When the flag is cleared, the MUX tasklet is woken up. However, as MUX iocb does not treat Rx for the moment, this is useless. It must be fix to prevent possible freeze on POST transfers. In practice, for the moment the HTX buffer is never full as the current Rx code is limited by the quic-conn receive buffer size and the incomplete flow-control implementation. So for now this patch is not testable under the current conditions.	2022-05-02 11:19:02 +02:00
Frédéric Lécaille	c40e19d711	BUG/MINOR: quic: Missing time threshold multiplifier for loss delay computation It seems this multiplier ended up in oblivion. Indeed a multiplier must be applied to the loss delay expressed as an RTT multiplier: 9/8. So, some packets were detected as lost too soon, leading to be retransmitted too early!	2022-04-29 16:46:56 +02:00
Frédéric Lécaille	1601395063	MINOR: quic: moving code for QUIC loss detection qc_qc_packet_loss_lookup() is definitively a QUIC loss detection function.	2022-04-29 16:46:56 +02:00
Frédéric Lécaille	88e5741c53	CLEANUP: quic: Remaining fprintf() debug trace Development remaining trace.	2022-04-29 16:46:56 +02:00
Frédéric Lécaille	1231d3c179	MINOR: quic: Drop 0-RTT packets without secrets If we received 0-RTT packets and no secrets were provided by the TLS stack we must drop them.	2022-04-29 16:46:56 +02:00
Amaury Denoyelle	74cf237ecd	MEDIUM: quic: do not ack packet with invalid STREAM If the MUX cannot handle immediately nor buffer a STREAM frame, the packet containing it must not be acknowledge. This is in conformance with the RFC9000. qcc_recv() return codes have been adjusted to differentiate an invalid frame with an already fully received offset which must be acknowledged.	2022-04-29 16:16:19 +02:00
Amaury Denoyelle	d46e335683	MEDIUM: quic: do not ACK packet with STREAM if MUX not present If a packet contains a STREAM frame but the MUX is not allocated, the frame cannot be enqueued. According to the RFC9000, we must not acknowledge the packet under this condition. This may prevents a bug with firefox which keeps trying on refreshing the web page. This issue has already been detected before closing state implementation : haproxy wasn't emitted CONNECTION_CLOSE and keeps acknowledge STREAM frames despite not handle them. In the future, it might be necessary to respond with a CONNECTION_CLOSE if the MUX has already been freed.	2022-04-29 16:15:47 +02:00
Willy Tarreau	4173f4ea29	BUG/MINOR: conn_stream: do not confirm a connection from the frontend path In issue #1468 it was reported that sometimes server-side connection attempts were only validated after the "timeout connect" value, and that would only happen with an H2 client. A long code analysis with the output dumps showed only one possible call path: an I/O event on the frontend while reading had just been disabled calls h2_wake() which in turns wakes cs_conn_io_cb(), which tries cs_conn_process() and cs_notify(), which sees that the other side is not blocked (already in CS_ST_CON) and tries cs_chk_snd() on it. But on that side the connection had just finished to be set up and not yet woken the stream up, cs_notify() would then call cs_conn_send() which succeeds and passes the connection to CS_ST_RDY. The problem is that nothing new happened on the frontend side so there's no reason to wake the stream up and the backend-side conn_stream remains in CS_ST_RDY state with the stream never being woken up. Once the "timeout connect" strikes, process_stream() is woken up and finds the connection finally setup, so it ignores the timeout and goes on. The number of conditions to meet to reproduce this is huge, which also explains why the reporter says it's "occasional" and we were never able to reproduce it in the lab. It needs at least reads to be disabled and immediately re-enabled on the frontend side (e.g. buffer full) with an I/O even reported before the poller had an opportunity to be disabled but with no subscribe being reinstalled, so that sock_conn_iocb() has no other choice but calling h2_wake(), and exactly at the same time the backend connection must finish to set up so that it was not yet reported by the poller, the data were sent and the polling for writes disabled. Several factors are to be considered here: - h2_wake() should probably not call h2_wake_some_streams() for ret >= 0 (common case), but only if some special event is reported for at least one stream; that part is sensitive though as in the past we managed to lose some rare cases (e.g. restart processing after a pause), and such wakeups are extremely rare so we'd better make that effort once in a while. - letting a lazy forward attempt on the frontend confirm a backend connection establishment is too smart to be reliable. That wasn't in fact the intent and it's inherited from the very old code where muxes didn't exist and where it was guaranteed that an even at this layer would wake everyone up. Here the best thing to do is to refrain from attempting to forward data until the connection is confirmed. This will let the poller report the connect() event to the backend side which will process it as it should and does in all other cases. Thanks to Jimmy Crutchfield for having reported useful traces and tested patches. This will have to be backported to all stable branches after some observation. Before 2.6 the function is stream_int_chk_snd_conn(), and the flag to remove is SI_SB_CON.	2022-04-29 15:32:14 +02:00
Christopher Faulet	0055d5693e	MINOR: httpclient: Don't use co_set_data() to decrement output The use of co_set_data() should be strictly limited to setting the amount of existing data to be transmitted. It ought not be used to decrement the output after the data have left the buffer, because doing so involves performing incorrect calculations using co_data() that still comprises data that are not in the buffer anymore. Let's use c_rew() for this, which is made exactly for this purpose, i.e. decrement c->output by as much as requested.	2022-04-29 14:12:42 +02:00
Christopher Faulet	6b4f1f64a8	BUG/MINOR: httpclient: Count metadata in size to transfer via htx_xfer_blks() When HTX blocks are transfer from the HTTP client context to the request channel, via htx_xfer_blks() function, the metadata must also be counted, in addition to the data size. Otherwise, expected payload size will not be copied because the metadata of an HTX block (8 bytes) will be reserved. And if the payload size is lower than 8 bytes, nothing will be copied. Thus only a zero-copy will be able to copy the payload. This issue is 2.6-specific, no backport is needed.	2022-04-29 14:12:42 +02:00
Christopher Faulet	534645d6c0	BUG/MEDIUM: httpclient: Fix loop consuming HTX blocks from the response channel When the HTTP client consumes the response, it loops on the HTX message to copy blocks content and it removes blocks by calling htx_remove_blk(). But this function removes a block and returns the next one in the HTX message. The result must be used instead of using htx_get_next(). It is especially important because the block used in htx_get_next() loop was removed. It only works because the message is not defragmented during the loop. In addition, the loop on the response was simplified to iter on blocks instead of positions. This patch must be backported to 2.5.	2022-04-29 14:12:42 +02:00
Christopher Faulet	a6c4a48341	BUG/MEDIUM: conn-stream: Don't erase endpoint flags on reset Only CS_EP_ERROR flag is now removed from the endpoint when a reset is performed. When a new the endpoint is allocated, flags are preserved. It is the caller responsibility to remove other flags, depending on its need. Concretly, during a connection retry or a L7 retry, we must preserve flags. In tcpcheck and the CLI, we reset flags. This patch is 2.6-specific. No backport needed.	2022-04-29 14:12:42 +02:00
William Lallemand	04994de642	BUG/MINOR: httpclient/ssl: use the correct verify constant The SSL_SERVER_VERIFY_* constants were incorrectly set on the httpclient server verify. The right constants are SSL_SOCK_VERIFY_* . This could cause issues when using "httpclient-ssl-verify" or when the SSL certificates can't be loaded. No backport needed	2022-04-28 19:35:21 +02:00
Frédéric Lécaille	3e26698f89	MINOR: quic: Drop 0-RTT packets if not allowed Drop the 0-RTT packets for a listener without early data configuration enabled.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	4646cf3b70	CLEANUP: quic: Rely on the packet length set by qc_lstnr_pkt_rcv() This function is used to parse the QUIC packets carried by a UDP datagram. When a correct packet could be found, the ->len RX packet structure value is set to the packet length value. On the contrary, it is set to the remaining number of bytes in the UDP datagram if no correct QUIC packet could be found. So, there is no need to make this function return a status value. It allows the caller to parse any QUIC packet carried by a UDP datagram without this.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	87373e7269	BUG/MINOR: quic: Missing Initial packet length check Any client Initial packet carried in a datagram smaller than QUIC_INITIAL_PACKET_MINLEN(200) bytes must be discarded. This does not mean we must discard the entire datagram. So we must at least try to parse the packet length before dropping the packet and return its length from qc_lstnr_pkt_rcv().	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	77cb38d22d	BUG/MEDIUM: quic: Possible crash on STREAM frame loss A crash is possible under such circumtances: - The congestion window is drastically reduced to its miniaml value when a quic listener is experiencing extreme packet loss ; - we enqueue several STREAM frames to be resent and some of them could not be transmitted ; - some of the still in flight are acknowledged and trigger the stream memory releasing ; - when we come back to send the remaing STREAM frames, haproxy crashes when it tries to build them. To fix this issue, we mark the STREAM frame as lost when detected as lost. Then a lookup if performed for the stream the STREAM frames are attached to before building them. They are released if the stream is no more available or the data range of the frame is consumed.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	dafbde6c8c	MINOR: quic: Wake up the mux to probe with new data When we have to probe the peer, we must first try to send new data. This is done here waking up the mux after having set the number of maximum number of datagrams to send to QUIC_MAX_NB_PTO_DGRAMS (2). Of course, this is only the case if the mux was subscribed to SEND events.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	d8b798d7ef	BUG/MINOR: quic: Traces fix about remaining frames upon packet build failure	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	834399c24a	BUG/MINOR: quic: Avoid sending useless PADDING frame This may happen in rare cases with extreme packet loss (30% for both TX and RX) which leads the congestion window to decrease down to its minimal value (two datagrams). Under such circumtances, no ack-eliciting frame can be added to a packet by qc_build_frms(). In this case we must cancel the packet building process if there is no ACK or probe (PING frame) to send.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	573b56b774	BUG/MINOR: quic: Wrong returned status by qc_build_frms() This function must return a successful status as soon as it could be build a frame to be embedded by a packet. This behavior was broken by the last modifications. This was due to a dangerous "ret = 1" statement inside a loop. This statement must be reach only if we go out of a switch/case after a "break" statement. Add comments to mention this information.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	337108ecda	MINOR: quic: Do not send ACK frames when probing When we are probing, we do not receive packets, furthermore all ACK frames have already been sent. This is useless to send ACK when probing the peer. This modification does not reset the flag which marks the connection as requiring an ACK frame to be sent. If this is the case, this will be taken into an account by after the probing process.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	7aef5f4c3f	MEDIUM: quic: Enable the new datagram probing process Make the two I/O handlers quic_conn_io_cb() and quic_conn_app_io_cb() call qc_dgrams_retransmit() after probing retransmissions need was detected by the timer task (qc_process_timer()). We must modify qc_prep_pkts() to support QUIC_TLS_ENC_LEVEL_NONE as <next_tel> parameter when called from qc_dgrams_retransmit().	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	da342556c3	MEDIUM: quic: Mark copies of acknowledged frames as acknowledged We call qc_release_frm() to do so from this function everywhere a frame is released.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	1809c33d6e	MINOR: quic: Mark packets as probing with old data When probing retranmissions with old data are needed for the connection we mark the packets as probing with old data to track them when acknowledged: we do not resend frames with old data when lost.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	3e3a621447	MINOR: quic: old data distinction for qc_send_app_pkt() Modify qc_send_app_pkt() to distinguish the case where it sends new data against the case where it sends old data during probing retransmissions. We add <old_data> boolean parameter to this function to do so. The mux never directly send old data when probing retransmissions are needed by the connection.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	96367158ab	MEDIUM: quic: qc_requeue_nacked_pkt_tx_frms() rework This function is used to requeue the TX frames from TX packets which have been detected as lost. The modifications consist in avoiding resending frames from duplicated frames used to probe the peer. This is useless. Only the original frames loss must be taken into an account because detected as lost before the retransmitted frames. If these latter are also detected as lost, other duplicated frames would have been retransmitted before their loss detection.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	e248e378c8	MEDIUM: quic: Retransmission functions rework qc_prep_fast_retrans() and qc_prep_hdshk_fast_retrans() are modified to take two list of frames as parameters. Two lists are needed for qc_prep_hdshk_fast_retrans() to build datagrams with two packets during handshake. qc_prep_fast_retrans() needs two lists of frames to be used to send two datagrams with one list by datagram.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	a9568411e4	MEDIUM: quic: New functions for probing rework We want to be able to resend frames from list of frames during handshakes to resend datagrams with the same frames as during the first transmissions. This leads to decrease drasctically the chances of frame fragmentation due to variable lengths of the packet fields. Furthermore the frames were not duplicated when retransmitted from a packet to another one. This must be the case only during packet loss dectection. qc_dup_pkt_frms() is there to duplicate the frames from an input list to an output list. A distinction is made between STREAM frames and the other ones because we can rely on the "acknowledged offset" the aim of which is to track the number of bytes which were acknowledged from TX STREAM frames. qc_release_frm() in addition to release the frame passed as parameter, also mark the duplicate STREAM frames as acknowledeged. qc_send_hdshk_pkts() is the qc_send_app_pkts() counterpart to send datagrams from at most two list of frames to be able to coalesced packets from two different packet number spaces qc_dgrams_retransmit() is there to probe the peer with datagrams depending on the need of the packet number spaces which must be flag with QUIC_FL_PKTNS_PROBE_NEEDED by the PTO timer task (qc_process_timer()).	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	3ef729a643	MINOR: quic: process_timer() rework Add QUIC_FL_CONN_RETRANS_NEEDED connection flag definition to mark a quic_conn struct as needing a retranmission. Add QUIC_FL_PKTNS_PROBE_NEEDED to mark a packet number space as needing a datagram probing. Set these flags from process_timer() to trigger datagram probings. Do not initiate anymore datagrams probing from any quic encryption level. This will be done from the I/O handlers (quic_conn_io_cb() during handshakes and quic_conn_app_io_cb() after handshakes).	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	e87b3ee9f5	MINOR: quic: Add traces about TX frame memory releasing Add such traces in qc_treat_acked_tx_frm(). This should be helpful to track memory leak issues for TX frames.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	b44cbc68a6	MINOR: quic: Do not retransmit frames from coalesced packets Add QUIC_FL_TX_PACKET_COALESCED flag to mark a TX packet as coalesced with others to build a datagram. Ensure we do not directly retransmit frames from such coalesced packets. They must be retransmitted from their packet number spaces to avoid duplications.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	b917191817	MINOR: quic: Prepare quic_frame struct duplication We want to track the frames which have been duplicated during retransmissions so that to avoid uselessly retransmitting frames which would already have been acknowledged. ->origin new member is there to store the frame from which a copy was done, ->reflist is a list to store the frames which are copies. Also ensure all the frames are zeroed and that their ->reflist list member is initialized. Add QUIC_FL_TX_FRAME_ACKED flag definition to mark a TX frame as acknowledged.	2022-04-28 16:22:40 +02:00
Frédéric Lécaille	fc88844d2c	MINOR: quic: Improve qc_prep_pkts() flexibility We want to be able to chosse the list of frames we want to prepare in packets to be send. This is to modify the retransmission process (to come).	2022-04-28 16:22:40 +02:00
Amaury Denoyelle	03cc62c840	MINOR: quic: decode as much STREAM as possible Add a loop in the bidi STREAM function. This will call repeatdly qcc_decode_qcs() and dequeue buffered frames. This is useful when reception of more data is interrupted because the MUX buffer was full. qcc_decode_qcs() has probably free some space so it is useful to immediatly retry reception of buffered frames of the qcs tree. This may fix occurences of stalled Rx transfers with large payload. Note however that there is still room for improvment. The conn-stream layer is not able at this moment to retrigger demuxing. This is because the mux io-handler does not treat Rx : this may continue to cause stalled tranfers.	2022-04-28 16:10:10 +02:00
Amaury Denoyelle	48f01bda86	MINOR: h3: support DATA demux if buffer full Previously, h3 layer was not able to demux a DATA frame if not fully received in the Rx buffer. This causes evident limitation and prevents to be able to demux a frame bigger than the buffer. Improve h3_data_to_htx() to support partial frame demuxing. The demux state is preserved in the h3s new fields : this is useful to keep the current type and length of the demuxed frame.	2022-04-28 15:44:19 +02:00
Amaury Denoyelle	67e92d3652	MINOR: h3: implement h3 stream context Define a new structure h3s used to provide context for a H3 stream. This structure is allocated and stored in the qcs thanks to previous commit which provides app-layer context storage. For now, h3s is empty. It will soon be completed to be able to support stateful demux : this is required to be able to demux an incomplete frame if the rx buffer is full.	2022-04-28 15:44:19 +02:00
Amaury Denoyelle	47447af1ef	MINOR: mux-quic: add a app-layer context in qcs Define 2 new callback for qcc_app_ops : attach and detach. They are called when a qcs instance is respectively allocated and freed. If implemented, they can allocate a custom context stored in the new abstract field ctx of qcs. For now, h3 and hq-interop does not use these new callbacks. They will be soon implemented by the h3 layer to allocate a context used for stateful demuxing. This change is required to support the demuxing of H3 frames bigger than a buffer.	2022-04-28 15:44:19 +02:00
Amaury Denoyelle	314578a54f	MINOR: h3: change frame demuxing API Edit the functions used for HEADERS and DATA parsing. They now return the number of bytes handled. This change will help to demux H3 frames bigger than the buffer.	2022-04-28 15:44:19 +02:00
Amaury Denoyelle	3df8ca0a4d	MINOR: mux-quic: partially copy Rx frame if almost full buf Improve the reception for STREAM frames. In qcc_recv(), if the frame is bigger than the remaining space in rx buffer, do not reject it wholly. Instead, copy as much data as possible. The rest of the data is buffered. This is necessary to handle H3 frames bigger than a buffer. The H3 code does not demux until the frame is complete or the buffer is full. Without this, the transfer on payload larger than the Rx buffer can rapidly freeze.	2022-04-28 15:42:21 +02:00
Amaury Denoyelle	30f23f53d2	BUG/MEDIUM: h3: fix use-after-free on mux Rx buffer wrapping Handle wrapping buffer in h3_data_to_htx(). If data is wrapping, first copy the contiguous data, then copy the data in front of the buffer. Note that h3_headers_to_htx() is not able to handle wrapping data. For the moment, a BUG_ON was added as a reminder. This cas never happened, most probably because HEADERS is the first frame of the stream.	2022-04-28 15:18:20 +02:00
Amaury Denoyelle	0fa14a69e8	BUG/MINOR: h3: fix incomplete POST requests Always set HTX flag HTX_SL_F_XFER_LEN for http/3. This is correct becuase the size of H3 requests is always known thanks to the protocol framing. This may fix occurences of incomplete POST requests when the client side of the connection has been closed before.	2022-04-28 15:14:25 +02:00
Amaury Denoyelle	44d0912f7b	MINOR: mux-quic: count local flow-control stream limit on reception Add new qcs fields to count the sum of bytes received for each stream. This is necessary to enforce flow-control for reception on the peer. For the moment, the implementation is partial. No MAX_STREAM_DATA or FLOW_CONTROL_ERROR are emitted. BUG_ON statements are here as a remainder. This means that for the moment we do not support POST payloads greater that the initial max-stream-data announced (256k currently). At least, we now ensure that we never buffer a frame which overflows the flow-control limit : this ensures that the memory consumption per stream should stay under control.	2022-04-28 14:56:14 +02:00
Amaury Denoyelle	17014a64bf	BUG/MINOR: mux-quic: fix leak if cs alloc failure qcs and its field are not properly freed if the conn-stream allocation fail in qcs_new(). Fix this by having a proper deinit code with a dedicated label.	2022-04-28 14:47:53 +02:00
Amaury Denoyelle	fe8f555f5c	MINOR: mux-quic: adjust comment on emission function Comments were not properly edited since the splitting of functions for stream emission. Also "payload" argument has been renamed to "in" as it better reflects the function purpose.	2022-04-28 14:47:21 +02:00
Amaury Denoyelle	b50f311c50	BUG/MINOR: mux-quic: fix build in release mode Fix build when not using DEBUG_STRICT. 'ret' is reported as unused as it is only tested in a BUG_ON statement.	2022-04-28 14:42:39 +02:00
Willy Tarreau	faafe4bf16	CLEANUP: connections/deinit: destroy the idle_conns tasks This adds a deinit_idle_conns() function that's called on deinit to release the per-thread idle connection management tasks. The global task was already taken care of.	2022-04-27 18:54:08 +02:00
Willy Tarreau	e01b08db6a	CLEANUP: listeners/deinit: release accept queue tasklets on deinit There was no function to release these ones, they were only created so the patch adds an accept_queue_deinit() call.	2022-04-27 18:42:47 +02:00
Willy Tarreau	226866e1bb	CLEANUP: deinit: release the config postparsers These ones were not released either, it just requires to export the list ("postparsers") and it makes valgrind happy.	2022-04-27 18:07:24 +02:00
Willy Tarreau	65009ebde1	CLEANUP: deinit: release the pre-check callbacks The freeing of pre-check callbacks was missing when this feature was recently added with commit `b53eb8790` ("MINOR: init: add the pre-check callback"), let's do it to make valgrind happy.	2022-04-27 18:02:54 +02:00
Willy Tarreau	d941146583	CLEANUP: chunks: release trash also in deinit Tim reported in issue #1676 that just like startup_logs, trash buffers are not released on deinit since they're thread-local, but valgrind notices it when quitting before creating threads like "-c -f ...". Let's just subscribe the function to deinit in addition to threads' end. The two "free(x);x=NULL;" in free_trash_buffers_per_thread() were also simplified using ha_free().	2022-04-27 17:55:41 +02:00
Willy Tarreau	032e700e8b	CLEANUP: errors: also call deinit_errors_buffers() on deinit() Tim reported in issue #1676 that we don't release startup logs if we warn during startup and quit before creating threads (e.g. -c -f ...). Let's subscribe deinit_errors_buffers() to both thread's end and deinit. That's OK since it uses both per-thread and global variables, and is idempotent.	2022-04-27 17:50:53 +02:00
Thomas Pr�ckl	10243938db	MINOR: ssl: add a new global option "tune.ssl.hard-maxrecord" Low footprint client machines may not have enough memory to download a complete 16KB TLS record at once. With the new option the maximum record size can be defined on the server side. Note: Before limiting the the record size on the server side, a client should consider using the TLS Maximum Fragment Length Negotiation Extension defined in RFC6066. This patch fixes GitHub issue #1679.	2022-04-27 16:53:43 +02:00
Willy Tarreau	243e68b552	BUG/MINOR: pools: make sure to also destroy shared pools in pool_destroy_all() In issue #1677, Tim reported that we don't correctly free some shared pools on exit. What happens in fact is that pool_destroy() is meant to be called once per pool pointer, as it decrements the use count for each pass and only releases the pool when it reaches zero. But since pool_destroy_all() iterates over the pools list, it visits each pool only once and will not eliminate some of them, which thus remain in the list. In an ideal case, the function should loop over all pools for as long as the list is not empty, but that's pointless as we know we're exiting, so let's just set the users count to 1 before the call so that pool_destroy() knows it can delete and release the entry. This could be backported to all versions (memory.c in 2.0 and older) but it's not a real problem in practice.	2022-04-27 11:33:13 +02:00
Willy Tarreau	974358954b	BUILD: fd: disguise the fd_set_nonblock/cloexec result We thought that we could get rid of some DISGUISE() with commit `a80e4a354` ("MINOR: fd: add functions to set O_NONBLOCK and FD_CLOEXEC") thanks to the calls being in a function but that was without counting on Coverity. Let's put it directly in the function since most if not all callers don't care about this result.	2022-04-27 10:52:21 +02:00
Tim Duesterhus	77b3db0fbd	MINOR: Call deinit_and_exit(0) for `haproxy -vv` It appears that it is safe to call perform a clean deinit at this point, so let's do this to exercise the deinit paths some more. Running `valgrind --leak-check=full --show-leak-kinds=all ./haproxy -vv` with this change reports: ==261864== HEAP SUMMARY: ==261864== in use at exit: 344 bytes in 11 blocks ==261864== total heap usage: 1,178 allocs, 1,167 frees, 1,102,089 bytes allocated ==261864== ==261864== 24 bytes in 1 blocks are still reachable in loss record 1 of 2 ==261864== at 0x483DD99: calloc (in /usr/lib/x86_64-linux-gnu/valgrind/vgpreload_memcheck-amd64-linux.so) ==261864== by 0x324BA6: hap_register_pre_check (init.c:92) ==261864== by 0x155824: main (haproxy.c:3024) ==261864== ==261864== 320 bytes in 10 blocks are still reachable in loss record 2 of 2 ==261864== at 0x483DD99: calloc (in /usr/lib/x86_64-linux-gnu/valgrind/vgpreload_memcheck-amd64-linux.so) ==261864== by 0x26E54E: cfg_register_postparser (cfgparse.c:4238) ==261864== by 0x155824: main (haproxy.c:3024) ==261864== ==261864== LEAK SUMMARY: ==261864== definitely lost: 0 bytes in 0 blocks ==261864== indirectly lost: 0 bytes in 0 blocks ==261864== possibly lost: 0 bytes in 0 blocks ==261864== still reachable: 344 bytes in 11 blocks ==261864== suppressed: 0 bytes in 0 blocks which is looking pretty good.	2022-04-27 05:01:27 +02:00
Tim Duesterhus	0b7031b37d	BUG/MINOR: resolvers: Fix memory leak in resolvers_deinit() A config like the following: global stats socket /run/haproxy/admin.sock mode 660 level admin expose-fd listeners resolvers unbound nameserver unbound 127.0.0.1:53 will report the following leak when running a configuration check: ==241882== 6,991 (6,952 direct, 39 indirect) bytes in 1 blocks are definitely lost in loss record 8 of 13 ==241882== at 0x483DD99: calloc (in /usr/lib/x86_64-linux-gnu/valgrind/vgpreload_memcheck-amd64-linux.so) ==241882== by 0x25938D: cfg_parse_resolvers (resolvers.c:3193) ==241882== by 0x26A1E8: readcfgfile (cfgparse.c:2171) ==241882== by 0x156D72: init (haproxy.c:2016) ==241882== by 0x156D72: main (haproxy.c:3037) because the `.px` member of `struct resolvers` is not freed. The offending allocation was introduced in `c943799c86` which is a reorganization that happened during development of 2.4.x. This fix can likely be backported without issue to 2.4+ and is likely not needed for earlier versions as the leak happens during deinit only.	2022-04-26 23:42:10 +02:00
Tim Duesterhus	2b7fa9d668	CLEANUP: Destroy `http_err_chunks` members during deinit To make the deinit function a proper inverse of the init function we need to free the `http_err_chunks`: ==252081== 311,296 bytes in 19 blocks are still reachable in loss record 50 of 50 ==252081== at 0x483B7F3: malloc (in /usr/lib/x86_64-linux-gnu/valgrind/vgpreload_memcheck-amd64-linux.so) ==252081== by 0x2727EE: http_str_to_htx (http_htx.c:914) ==252081== by 0x272E60: http_htx_init (http_htx.c:1059) ==252081== by 0x26AC87: check_config_validity (cfgparse.c:4170) ==252081== by 0x155DFE: init (haproxy.c:2120) ==252081== by 0x155DFE: main (haproxy.c:3037)	2022-04-26 23:39:43 +02:00
Christopher Faulet	eab175771d	BUG/MEDIUM: http-ana: Fix memleak in redirect rules with ignore-empty option A memory leak was introduced when ignore-empty option was added to redirect rules. If there is no location, when this option is set, the redirection is aborted and the processing continues. But when this happened, the trash buffer allocated to format the redirect response was not released. The bug was introduced by commit `bc1223be7` ("MINOR: http-rules: add a new "ignore-empty" option to redirects."). This patch should fix the issue #1675. It must be backported to 2.5.	2022-04-26 20:47:22 +02:00
Remi Tricot-Le Breton	4d7fdc65d4	MINOR: connection: Add way to disable active connection closing during soft-stop If the "close-spread-time" option is set to "infinite", active connection closing during a soft-stop can be disabled. The 'connection: close' header or the GOAWAY frame will not be added anymore to the server's response and active connections will only be closed once the clients disconnect. Idle connections will not be closed all at once when the soft-stop starts anymore, and each idle connection will follow its own timeout based on the multiple timeouts set in the configuration (as is the case during regular execution). This feature request was described in GitHub issue #1614. This patch should be backported to 2.5. It depends on 'MEDIUM: global: Add a "close-spread-time" option to spread soft-stop on time window'.	2022-04-26 19:56:47 +02:00
William Lallemand	03a32e5dd2	BUG/MEDIUM: ssl/cli: fix yielding in show_cafile_detail HAProxy crashes when "show ssl ca-file" is being called on a ca-file which contains a lot of certificates. (127 in our test with /etc/ssl/certs/ca-certificates.crt). The root cause is the fonction does not yield when there is no available space when writing the details, and we could write a lot. To fix the issue, we try to put the data with ci_putchk() after every show_cert_detail() and we yield if the ci_putchk() fails. This patch also cleans up a little bit the code: - the end label is now a real end with a return 1; - i0 is used instead of (long)p1 - the ID is stored upon yield	2022-04-26 19:35:43 +02:00
William Lallemand	f1344b3cee	MEDIUM: httpclient: re-enable the verify by default Since the httpclient verify now has a fallback which disable the SSL in the httpclient without exiting haproxy at startup, we can safely re-enable it by default. It could still be disabled with "httpclient-ssl-verify none".	2022-04-26 16:15:23 +02:00
William Lallemand	4cfbf3c014	BUG/MINOR: ssl: memory leak when trying to load a directory with ca-file This patch fixes a memory leak of the ca structure when trying to load a directory with the ca-file directive. No backport needed.	2022-04-26 16:15:23 +02:00
William Lallemand	b0c4827c2f	BUG/MINOR: ssl: free the cafile entries on deinit The cafile_tree was never free upon deinit, making valgrind and ASAN complains when haproxy quits. This could be backported as far as 2.2 but it requires the ssl_store_delete_cafile_entry() helper from `5daff3c8ab`.	2022-04-26 16:15:23 +02:00
William Lallemand	c2d3db44ad	BUG/MINOR: httpclient/lua: error when the httpclient_start() fails Jump to an luaL_error() when the httpclient fails, it could be the result of an allocation failure, or even a wrong URL. Must be backported in 2.5.	2022-04-26 16:15:23 +02:00
William Lallemand	4006b0f130	MEDIUM: httpclient: disable SSL when the ca-file couldn't be loaded Emit a warning when the ca-file couldn't be loaded for the httpclient, and disable the SSL of the httpclient. We must never be in a case where the verify is disabled without any configuration, so better disable the SSL completely. Move the check on the scheme above the initialization of the applet so we could abort before initializing the appctx.	2022-04-26 16:15:23 +02:00
Willy Tarreau	7e2e4f8401	CLEANUP: tree-wide: remove 25 occurrences of unneeded fcntl.h There were plenty of leftovers from old code that were never removed and that are not needed at all since these files do not use any definition depending on fcntl.h, let's drop them.	2022-04-26 10:59:48 +02:00
Willy Tarreau	382474348c	CLEANUP: tree-wide: use fd_set_nonblock() and fd_set_cloexec() This gets rid of most open-coded fcntl() calls, some of which were passed through DISGUISE() to avoid a useless test. The FD_CLOEXEC was most often set without preserving previous flags, which could become a problem once new flags are created. Now this will not happen anymore.	2022-04-26 10:59:48 +02:00
Willy Tarreau	a80e4a3546	MINOR: fd: add functions to set O_NONBLOCK and FD_CLOEXEC Instead of seeing each location manipulate the fcntl() themselves and often forget to check previous flags, let's centralize the functions to do this. It also allows to drop fcntl.h from most call places and will ease the adoption of different OS-specific mechanisms if needed. Note that the fd_set_nonblock() function purposely doesn't check the previous flags as it's meant to be used on new FDs only.	2022-04-26 10:59:48 +02:00
Remi Tricot-Le Breton	b4f5fac886	BUG/MINOR: connection: "connection:close" header added despite 'close-spread-time' Despite what the 'close-spread-time' option should do, the 'connection:close' header was always added to HTTP responses during soft-stops even with a soft-stop window defined. This patch adds the proper random based closing to HTTP connections during a soft-stop (based on the time left in the soft close window). It should be backported to 2.5 once 'MEDIUM: global: Add a "close-spread-time" option to spread soft-stop on time window' is backported as well.	2022-04-26 10:50:47 +02:00
Willy Tarreau	acef5e27b0	MINOR: tree-wide: always consider EWOULDBLOCK in addition to EAGAIN Some older systems may routinely return EWOULDBLOCK for some syscalls while we tend to check only for EAGAIN nowadays. Modern systems define EWOULDBLOCK as EAGAIN so that solves it, but on a few older ones (AIX, VMS etc) both are different, and for portability we'd need to test for both or we never know if we risk to confuse some status codes with plain errors. There were few entries, the most annoying ones are the switch/case because they require to only add the entry when it differs, but the other ones are really trivial.	2022-04-25 20:32:15 +02:00
Willy Tarreau	197715ae21	CLEANUP: compression: move the default setting of maxzlibmem to defaults __comp_fetch_init() only presets the maxzlibmem, and only when both USE_ZLIB and DEFAULT_MAXZLIBMEM are set. The intent is to preset a default value to protect the system against excessive memory usage when no setting is set by the user. Nowadays the entry in the global struct is always there so there's no point anymore in passing via a constructor to possibly set this value. Let's go the cleaner way by always presetting DEFAULT_MAXZLIBMEM to 0 in defaults.h unless these conditions are met, and always assigning it instead of pre-setting the entry to zero. This is more straightforward and removes some ifdefs and the last constructor. In addition, now the setting has a chance of being found.	2022-04-25 19:42:43 +02:00
Willy Tarreau	ebab60279e	BUILD: http: remove the two unused constructors in rules and ana __http_protocol_init() and __http_rules_init() were empty leftovers from a previous more intense use of constructors. Let's just get rid of them now.	2022-04-25 19:26:26 +02:00
Willy Tarreau	8ead1d084a	BUILD: thread: use initcall instead of a constructor The constructor present there could be replaced with an initcall. This one is set at level STG_PREPARE because it also zeroes the lock_stats, and it's a bit odd that it could possibly have been scheduled to run after other constructors that might already preset some of these locks by accident.	2022-04-25 19:23:17 +02:00
Willy Tarreau	79367f9a8d	BUILD: xprt: use an initcall to register the transport layers Transport layers (raw_sock, ssl_sock, xprt_handshake and xprt_quic) were using 4 constructors and 2 destructors. The 4 constructors were replaced with INITCALL and the destructors with REGISTER_POST_DEINIT() so that we do not depend on this anymore.	2022-04-25 19:18:24 +02:00
Willy Tarreau	740d749d77	BUILD: pollers: use an initcall to register the pollers Pollers are among the few remaining blocks still using constructors to register themselves. That's not needed anymore since the initcalls so better turn to initcalls.	2022-04-25 19:00:55 +02:00
Willy Tarreau	2df1fbf816	MINOR: init: add global setting "fd-hard-limit" to bound system limits On some systems, the hard limit for ulimit -n may be huge, in the order of 1 billion, and using this to automatically compute maxconn doesn't work as it requires way too much memory. Users tend to hard-code maxconn but that's not convenient to manage deployments on heterogenous systems, nor when porting configs to developers' machines. The ulimit-n parameter doesn't work either because it forces the limit. What most users seem to want (and it makes sense) is to respect the system imposed limits up to a certain value and cap this value. This is exactly what fd-hard-limit does. This addresses github issue #1622.	2022-04-25 18:04:49 +02:00
Willy Tarreau	7c9a0fe2a6	MEDIUM: backend: add new "balance hash <expr>" algorithm Almost all of our hash-based LB algorithms are implemented as special cases of something that can now be achieved using sample expressions, and some of them have adopted some options to adapt their behavior in ways that could also be achieved using converters. There are users who want to hash other parameters that are combined into variables, and who set headers from these values and use "balance hdr(name)" for this. Instead of constantly implementing specific options and having users hack around when they want a real hash, let's implement a native hash mode that applies to a standard sample expression. This way, any fetchable element (including variables) may be used to construct the hash, even modified by any converter if desired.	2022-04-25 16:09:26 +02:00
Willy Tarreau	b9f30f398b	MINOR: sample: make the bool type cast to bin Any type except bool could cast to bin, while it can cast to string. That's a bit inconsistent, and prevents a boolean from being used as the entry of a hash function while any other type can. This is a problem when passing via variable where someone could use: ... set-var(txn.bar) always_false to temporarily disable something, but this would result in an empty hash output when later doing: ... var(txn.bar),sdbm Instead of using c_int2bin() as is done for the string output, better enfore an set of inputs or exactly 0 or 1 so that a poorly written sample fetch function does not result in a difficult to debug hash output.	2022-04-25 16:09:26 +02:00
Willy Tarreau	3705deea99	MINOR: sample: don't needlessly call c_none() in sample_fetch_as_type() Surprisingly, while about all calls to a sample cast function carefully avoid calling c_none(), sample_fetch_as_type() makes no effort regarding this, while by nature, the function is most often called with an expected output type similar to the one of the expresison. Let's add it to shorten the most common call path.	2022-04-25 16:09:26 +02:00
Willy Tarreau	1a8636d017	BUG/MINOR: sample: add missing use_backend/use-server contexts in smp_resolve_args The use_backend and use-server contexts were not enumerated in smp_resolve_args, and while use-server doesn't currently take an expression, at least use_backend supports that, and both entries ought to be listed for completeness. Now an error in a use_backend rule becomes more precise, from: [ALERT] (12373) : config : parsing [use-srv.cfg:33]: unable to find backend 'foo' referenced in arg 1 of sample fetch keyword 'nbsrv' in proxy 'echo'. to: [ALERT] (12307) : config : parsing [use-srv.cfg:33]: unable to find backend 'foo' referenced in arg 1 of sample fetch keyword 'nbsrv' in use_backend expression in proxy 'echo'. This may be backported though this is totally harmless.	2022-04-25 16:09:26 +02:00
Willy Tarreau	16daaf319c	BUG/MINOR: http-act: make release_http_redir() more robust Since commit `dd7e6c6dc` ("BUG/MINOR: http-rules: completely free incorrect TCP rules on error") free_act_rule() is called on some error paths, and one of them involves incomplete redirect rules that may cause a crash if the rule wasn't yet initialized, as shown in this config snippet: frontend ft mode http bind *:8001 http-request redirect location /%[always_false,sdbm] Let's simply make release_http_redir() more robust against null redirect rules. No backport needed since it seems that the only way to trigger this was the extra check above that was merged during 2.6-dev.	2022-04-25 16:09:26 +02:00
Christopher Faulet	643f1b7bef	BUG/MINOR: rules: Fix check_capture() function to use the right rule arguments The function checking captures defined in tcp-request content ruleset didn't use the right rule arguments. "arg.trk_ctr" was used instead of "arg.cap". This patch must be backported as far as 2.2.	2022-04-25 15:28:21 +02:00
Christopher Faulet	5796228aba	BUG/MEDIUM: rules: Be able to use captures defined in defaults section Since the 2.5, it is possible to define TCP/HTTP ruleset in defaults sections. However, rules defining a capture in defaults sections was not properly handled because they was not shared with the proxies inheriting from the defaults section. This led to crash when haproxy tried to store a new capture. So now, to fix the issue, when a new proxy is created, the list of captures points to the list of its defaults section. It may be NULL or not. All new caputres are prepended to this list. It is not a problem to share the same defaults section between several proxies, because it is not altered and we take care to not release it when corresponding proxies are freed but only when defaults proxies are freed. To do so, defaults proxies are now unreferenced at the end of free_proxy() function instead of the beginning. This patch should fix the issue #1674. It must be backported to 2.5.	2022-04-25 15:28:21 +02:00
Christopher Faulet	6c10f5c7bc	BUG/MINOR: rules: Forbid captures in defaults section if used by a backend Captures must only be defined in proxies with the frontend capabilities or in defaults sections used by proxies with the frontend capabilities. Thus, an extra check is added to be sure a defaults section defining a capture will never be references by a backend. Note that in this case, only named captures in "tcp-request content" or "http-request" rules are possible. It is not possible in a defaults section to decalre a capture slot. Not yet at least. This patch must be backported to 2.5. It is releated to issue #1674.	2022-04-25 15:05:04 +02:00
Amaury Denoyelle	7586bef6d7	BUG/MINOR: quic: fix use-after-free with trace on ACK consume When using qc_stream_desc_ack(), the stream instance may be freed if there is no more data in its buffers. This also means that all frames still stored waiting for ACK for this stream are freed via qc_stream_desc_free(). This is particularly important in quic_stream_try_to_consume() where we loop over the frames tree of the stream. A use-after-free is present in cas the stream has been freed in the trace "stream consumed" which dereference the frame. Fix this by first checking if the stream has been freed or not. This bug was detected by using ASAN + quic traces enabled.	2022-04-25 15:01:53 +02:00
Willy Tarreau	27fab1dcbc	MEDIUM: queue: use tasklet_instant_wakeup() to wake tasks It's long been known that queues didn't scale with threads for various reasons ranging from the cost of the queue lock to the cost of the massive amount of inter-thread wakeups. But some recent reports showing deplorable perfs with threads used at 100% CPU helped us notice that the two elements above add on top of each other: - with plenty of inter-thread wakeups, the scheduler takes a lot of time to dequeue pending tasks from the shared queue ; - the lock held by the scheduler to do this slows down subsequent task_wakeup() calls from the the queue that are made under the queue's lock - the queue's lock slows down addition of new requests to the queue and adds up to the number of needed queue entries for a steady traffic. But the cost of the share queue has no reason for being paid because it had already been paid when process_stream() added the request to the queue. As such an instant wakeup is perfectly fit for this. This is exactly what this patch does, it uses tasklet_instant_wakeup() to dequeue pending requests, which has the effect of not bloating the shared queue, hence not requiring the global queue lock, which in turn results in the wakeup to be much faster, and the queue lock to be much shorter. In the end, a test with 4k concurrent connections that was being limited to 40-80k requests/s before with 16 threads, some of which were stuck at 100% CPU now reaches 570k req/s with 4% idle. Given that it's been found that it was possible to trigger the watchdog on the queue lock under extreme conditions, and that such conditions could happen when users want to protect their servers during a DoS, it would definitely make sense to backport it to the most recent releases (2.5 and 2.4 seem like good candidates especially because their scheduler is modern enough to receive the change above). If a backport is performed, the following patch is needed: MINOR: task: add a new task_instant_wakeup() function	2022-04-22 19:11:59 +02:00
Amaury Denoyelle	f6df6b440a	BUG/MINOR: mux-quic: fix POST with abortonclose Remove CS_EP_EOS set erroneously on qc_rcv_buf(). This fixes POST with abortonclose. Previously, request was preemptively aborted by haproxy due to the incorrect EOS flag. For the moment, EOS flag is not set anymore. It should be set to warn about a premature close from the client.	2022-04-22 18:20:28 +02:00
Amaury Denoyelle	b710415e75	BUG/MEDIUM: mux-quic: fix stalled POST requets Add a missing notify RECV for conn-stream in qcc_decode_qcs(). This fixes stalled POST requests which may occur with some clients such as ngtcp2.	2022-04-22 18:18:34 +02:00
Christopher Faulet	7ae48a70d6	BUG/MAJOR: connection: Never remove connection from idle lists outside the lock Since the idle connections management changed to use eb-trees instead of MT lists, a lock must be acquired to manipulate servers idle/safe/available connection lists. However, it remains an unprotected use in connect_server(), when a connection is removed from an idle list if the mux has no more streams available. Thus it is possible to remove a connection from an idle list on a thread, while another one is looking for a idle connection. Of couse, this may lead to a crash. To fix the bug, we must take care to acquire the idle connections lock first. The bug was introduced by the commit `f232cb3e9` ("MEDIUM: connection: replace idle conn lists by eb trees"). The patch must be backported as far as 2.4.	2022-04-22 18:16:06 +02:00
William Lallemand	eaa703ef06	MEDIUM: httpclient/ssl: verify is configurable and disabled by default Disable temporary the SSL verify by default in the httpclient. The initialization of the @system-ca during the init of the httpclient is a problem in some cases. The verify can be reactivated with "httpclient-ssl-verify required" in the global section.	2022-04-22 18:05:17 +02:00
William Lallemand	c6ceba3170	MINOR: httpclient/mworker: disable in the master process Disable the httpclient in the master process.	2022-04-22 16:49:53 +02:00
William Lallemand	cf5cb0b524	MEDIUM: httpclient/ssl: verify required The httpclient HTTPS requests now enable the "verify required" option. To achieve this, the "@system-ca" ca-file is configured in the httpclient ssl server. Which means all the system CAs will be loaded at haproxy startup.	2022-04-22 15:45:47 +02:00
William Lallemand	2c8b0842bb	MEDIUM: httpclient: change the init sequence Change the init order of the httpclient, a different init sequence is required to allow a more complicated init. The init is splitted in two parts: - the first part is executed before config_check_validity(), which allows to create proxy and more advanced stuff than STG_INIT, because we might want to use stuff already initialized in haproxy (trash buffers for example) - the second part is executed after the config_check_validity(), currently it is used for the log configuration.	2022-04-22 15:45:47 +02:00
William Lallemand	b53eb8790e	MINOR: init: add the pre-check callback This adds a call to function <fct> to the list of functions to be called at the step just before the configuration validity checks. This is useful when you need to create things like it would have been done during the configuration parsing and where the initialization should continue in the configuration check. It could be used for example to generate a proxy with multiple servers using the configuration parser itself. At this step the trash buffers are allocated. Threads are not yet started so no protection is required. The function is expected to return non-zero on success, or zero on failure. A failure will make the process emit a succinct error message and immediately exit.	2022-04-22 15:45:47 +02:00
Christopher Faulet	eb50c01fef	MINOR: conn-stream: Make cs_detach_* private and use cs_destroy() from outside A conn-stream is never detached from an endpoint or an application alone, except on a reset. Thus, to avoid any error, these functions are now private. And cs_destroy() function is added to destroy a conn-stream. This function is called when a stream is released, on the front and back conn-streams, and when a health-check is finished.	2022-04-22 14:32:30 +02:00
Christopher Faulet	c6b60f1e00	MINOR: stream: Don't needlessly detach server endpoint on early client abort When a client abort is detected with the server conn-stream in CS_ST_INI state, there is no reason to detach the endpoing because we know there is no endpoint attached to this conn-stream. This patch depends on the commit "BUG/MEDIUM: conn-stream: Set back CS to RDY state when the appctx is created".	2022-04-22 14:32:30 +02:00
Christopher Faulet	a33ff7a8a7	BUG/MEDIUM: conn-stream: Set back CS to RDY state when the appctx is created When an appctx is created on the server side, we now set the corresponding conn-stream to ready state (CS_ST_RDY). When it happens, the backend conn-stream is in CS_ST_INI state. It is not consistant to let the conn-stream in this state because it means it is possible to have a target installed in CS_ST_INI state, while with a connection, the conn-stream is switch to CS_ST_RDY or CS_ST_EST state. It is especially anbiguous because we may be tempted to think there is no endpoint attached to the conn-stream before the CS_ST_CON state. And it is indeed the reason for a bug leading to a crash because a cs_detach_endp() is performed if an abort is detected on the backend conn-stream in CS_ST_INI state. With a mux or a appctx attached to the conn-stream, "->endp" field is set to NULL. It is unexpected. The API will be changed to be sure it is not possible. But it exposes a consistency issue with applets. So, the conn-stream must not stay in CS_ST_INI state when an appctx is attached. But there is no reason to set it in CS_ST_REQ. The conn-stream must be set to CS_ST_RDY to handle applets and connections in the same way. Note that if only the target is set but no appctx is created, the backend conn-stream is switched from CS_ST_INI to CS_ST_REQ state to be able to create the corresponding appctx. This part is unchanged. This patch depends on the commit "MINOR: backend: Don't allow to change backend applet". The ambiguity exists on previous versions. But the issue is 2.6-specific. Thus, no backport is needed.	2022-04-22 14:32:30 +02:00
Christopher Faulet	bb5b62ee5c	BUG/MINOR: backend: Don't allow to change backend applet This part was inherited from haproxy-1.5. But since a while (at least 1.8), the backend applet, once created, is no longer changed. Thus there is no reason to still check if the target has changed. And in fact, if it was still possible, there would be a memory leak because the old applet would be lost and never released. There is no reason to backport this fix because the leak only exists on a dead code path.	2022-04-22 14:14:27 +02:00
Christopher Faulet	1d216c7ec1	BUG/MINOR: cache: Disable cache if applet creation fails When we want to serve a resource from the cache, if the applet creation fails, the "cache-use" action must not yield. Otherwise, the stream will hang. Instead, we now disable the cache. Thus the request may be served by the server. This patch must be backported as far as 1.8.	2022-04-22 14:14:27 +02:00
Christopher Faulet	02ef0ff061	MINOR: conn-stream: Rely on endpoint shutdown flags to shutdown an applet cs_applet_shut() now relies on CS_EP_SH* flags to performed the applet shutdown. It means the applet release callback is called if there is no CS_EP_SHR or CS_EP_SHW flags set. And it set these flags, CS_EP_SHRR and CS_EP_SHWN more specifically, before exiting. This way, cs_applet_shut() is the really equivalent to cs_conn_shut().	2022-04-22 14:14:27 +02:00
Christopher Faulet	ca6c9bba82	CLEANUP: conn-stream: Rename cs_applet_release() This function does not release the applet but only call the applet release callback. It is equivalent to cs_conn_shut() but for applets. Thus the function is renamed cs_applet_shut().	2022-04-22 14:14:27 +02:00
Christopher Faulet	ff022a2b8c	CLEANUP: conn-stream: Rename cs_conn_close() and cs_conn_drain_and_close() These functions don't close the connection but only perform shutdown for reads and writes at the mux level. It is a bit ambiguous. Thus, cs_conn_close() is renamed cs_conn_shut() and cs_conn_drain_and_close() is renamed cs_conn_drain_and_shut(). These both functions rely on cs_conn_shutw() and cs_conn_shutr().	2022-04-22 14:14:27 +02:00
Christopher Faulet	0264212ba3	DEV: stream: Fix conn-streams dump in full stream message Since the recent changes about the conn-streams, the stream dump in "show sess all" command is a bit mangled. front and back conn-stream are now properly displayed (csf and csb). In addition, when there is no backend endpoint, "APPCTX" was always reported. Now, "NONE" is reported in this case. It is 2.6-specific. No backport needed.	2022-04-22 14:14:27 +02:00
Amaury Denoyelle	3eb892fd65	BUG/MINOR: mux-quic: remove dead code in qcs_xfer_data() Since previous patch MINOR: mux-quic: split xfer and STREAM frames build there is no way to report an error in qcs_xfer_data(). This should fix github issue #1669.	2022-04-22 09:51:56 +02:00
Willy Tarreau	e1e9f6bbe5	BUG/MEDIUM: logs: fix http-client's log srv initialization As anticipated in commit `211ea252d` ("BUG/MINOR: logs: fix logsrv leaks on clean exit"), there were indeed other corner cases that were not properly covered. Setting the http client's ring_name to NULL make the sink lookup crash on startup in sink_find () with a config as simple as: global log ring@buf0 local0 The fields must be properly initialized (both config file name and the ring_name). This only needs to be backported if/when the commit above is backported.	2022-04-22 09:40:44 +02:00
Amaury Denoyelle	a3daaec5a6	BUG/MINOR: mux-quic: handle null timeout Do not initialize mux task timeout if timeout client is set to 0 in the configuration. Check for the task before queuing it in qc_io_cb() or qc_detach(). This fix a crash when timeout client is 0 or undefined.	2022-04-21 16:35:46 +02:00
Amaury Denoyelle	f3e03a4066	BUG/MINOR: mux-quic: unsubscribe on release Unsubscribe from lower layer on qc_release. This ensures that the lower layer won't wake up a null tasklet after the MUX has been released and may prevent a crash.	2022-04-21 16:34:15 +02:00
Frédéric Lécaille	89a2ceb1fb	BUG/MEDIUM: quic: Possible crash with released mux It is possible the xprt layer have to process retransmitted STREAM frames after the mux was released. In this case, there is no need to try to wake it up.	2022-04-21 15:27:33 +02:00
Remi Tricot-Le Breton	f87c67e5e4	MINOR: ssl: Add 'show ssl providers' cli command and providers list in -vv option Starting from OpenSSLv3, providers are at the core of cryptography functions. Depending on the provider used, the way the SSL functionalities work could change. This new 'show ssl providers' CLI command allows to show what providers were loaded by the SSL library. This is required because the provider configuration is exclusively done in the OpenSSL configuration file (/usr/local/ssl/openssl.cnf for instance). A new line is also added to the 'haproxy -vv' output containing the same information.	2022-04-21 14:54:45 +02:00
Amaury Denoyelle	97e84c6c69	MINOR: cfg-quic: define tune.quic.conn-buf-limit Add a new global configuration option to set the limit of buffers per QUIC connection. By default, this value is set to 30.	2022-04-21 12:04:04 +02:00
Amaury Denoyelle	1b2dba531d	MINOR: mux-quic: implement immediate send retry Complete qc_send function. After having processed each qcs emission, it will now retry send on qcs where transfer can continue. This is useful when qc_stream_desc buffer is full and there is still data present in qcs buf. To implement this, each eligible qcs is inserted in a new list <qcc.send_retry_list>. This is done on send notification from the transport layer through qcc_streams_sent_done(). Retry emission until send_retry_list is empty or the transport layer cannot proceed more data. Several send operations are now called on two different places. Thus a new _qc_send_qcs() function is defined to factorize the code. This change should maximize the throughput during QUIC transfers.	2022-04-21 12:04:04 +02:00
Amaury Denoyelle	d2f80a2e63	MINOR: quic: limit total stream buffers per connection MUX streams can now allocate multiple buffers for sending. quic-conn is responsible to limit the total count of allowed allocated buffers. A counter is stored in the new field <stream_buf_count>. For the moment, the value is hardcoded to 30. On stream buffer allocation failure, the qcc MUX is flagged with QC_CF_CONN_FULL. The MUX is then woken up as soon as a buffer is freed, most notably on ACK reception.	2022-04-21 12:04:04 +02:00
Amaury Denoyelle	1b81dda3e0	MINOR: quic-stream: refactor ack management Acknowledge of STREAM has been complexified with the introduction of stream multi buffers. Two functions are executing roughly the same set of instructions in xprt_quic.c. To simplify this, move the code complexity in a new function qc_stream_desc_ack(). It will handle offset calculation, removal of data, freeing oldest buffer and freeing stream instance if required. The qc_stream_desc API is cleaner as qc_stream_desc_free_buf() ambiguous function has been removed.	2022-04-21 12:04:04 +02:00
Amaury Denoyelle	a456920491	MEDIUM: quic: implement multi-buffered Tx streams Complete the qc_stream_desc type to support multiple buffers on emission. The main objective is to increase the transfer throughput. The MUX is now able to transfer more data without having to wait ACKs. To implement this feature, a new type qc_stream_buf is declared. it encapsulates a buffer with a list element. New functions are defined to retrieve the current buffer, release it or allocate a new one. Each buffer is kept in the qc_stream_desc list until all of its data is acknowledged. On the MUX side, a qcs uses the current stream buffer to transfer data. Once the buffer is full, it is released and a new one will be allocated on a future qc_send() invocation.	2022-04-21 12:03:20 +02:00
Amaury Denoyelle	b22c0460d6	MINOR: quic-stream: add qc field Add a new member <qc> in qc_stream_desc structure. This change is possible since previous patch which add quic-conn argument to qc_stream_desc_new(). The purpose of this change is to simplify the future evolution of qc-stream-desc API. This will avoid to repeat qc as argument in various functions which already used a qc_stream_desc.	2022-04-21 11:55:29 +02:00
Amaury Denoyelle	e4301da5ed	MINOR: quic-stream: use distinct tree nodes for quic stream and qcs Simplify the model qcs/qc_stream_desc. Each types has now its own tree node, stored respectively in qcc and quic-conn trees. It is still necessary to mark the stream as detached by the MUX once all data is transfered to the lower layer. This might improve slightly the performance on ACK management as now only the lookup in quic-conn is necessary. On the other hand, memory size of qcs structure is increased.	2022-04-21 11:05:58 +02:00
Amaury Denoyelle	0cc02a345b	REORG: quic: use a dedicated module for qc_stream_desc Regroup all type definitions and functions related to qc_stream_desc in the source file src/quic_stream.c. qc_stream_desc complexity will be increased with the development of Tx multi-buffers. Having a dedicated module is useful to mix it with pure transport/quic-conn code.	2022-04-21 11:05:27 +02:00
Amaury Denoyelle	da6ad2092a	MINOR: mux-quic: split xfer and STREAM frames build Split qcs_push_frame() in two functions. The first one is qcs_xfer_data(). Its purpose is to transfer data from qcs.tx.buf to qc_stream_desc buffer. The second function is named qcs_build_stream_frm(). It generates a STREAM frame using qc_stream_desc buffer as payload. The trace events previously associated with qcs_push_frame() has also been split in two to reflect the new code structure. The purpose of this refactoring is first to better reflect how sending is implemented. It will also simplify the implementation of Tx multi-buffer per streams.	2022-04-21 09:27:43 +02:00
Remi Tricot-Le Breton	c69be7cd3c	BUILD: ssl: Fix compilation with OpenSSL 1.0.2 The DH parameters used for OpenSSL versions 1.1.1 and earlier where changed. For OpenSSL 1.0.2 and LibreSSL the newly introduced ssl_get_dh_by_nid function is not used since we keep the original parameters.	2022-04-20 22:34:44 +02:00
Remi Tricot-Le Breton	1d6338ea96	MEDIUM: ssl: Disable DHE ciphers by default DHE ciphers do not present a security risk if the key is big enough but they are slow and mostly obsoleted by ECDHE. This patch removes any default DH parameters. This will effectively disable all DHE ciphers unless a global ssl-dh-param-file is defined, or tune.ssl.default-dh-param is set, or a frontend has DH parameters included in its PEM certificate. In this latter case, only the frontends that have DH parameters will have DHE ciphers enabled. Adding explicitely a DHE ciphers in a "bind" line will not be enough to actually enable DHE. We would still need to know which DH parameters to use so one of the three conditions described above must be met. This request was described in GitHub issue #1604.	2022-04-20 17:30:55 +02:00
Remi Tricot-Le Breton	528b3fd9be	MINOR: ssl: Use DH parameters defined in RFC7919 instead of hard coded ones RFC7919 defined sets of DH parameters supposedly strong enough to be used safely. We will then use them when we can instead of our hard coded ones (namely the ffdhe2048 and ffdhe4096 named groups). The ffdhe2048 and ffdhe4096 named groups were integrated in OpenSSL starting with version 1.1.1. Instead of duplicating those parameters in haproxy for older versions of OpenSSL, we will keep using our own parameters when they are not provided by the SSL library. We will also need to keep our 1024 bits DH parameters since they are considered not safe enough to have a dedicated named group in RFC7919 but we must still keep it for retrocompatibility with old Java clients. This request was described in GitHub issue #1604.	2022-04-20 17:30:52 +02:00
Willy Tarreau	43041aaefd	BUILD: calltrace: fix wrong include when building with TRACE=1 calltrace wasn't updated after the move of "now" from time.h to clock.h. This must be backported to 2.5 where the breakage happened.	2022-04-19 08:23:30 +02:00
David CARLIER	7747d465d5	MINOR: tcp_sample: extend support for get_tcp_info to macOs. MacOS can feed fc_rtt, fc_rttvar, fc_sacked, fc_lost and fc_retrans so let's expose them on this platform. Note that at the tcp(7) level, the API is slightly different, as struct tcp_info is called tcp_connection_info and TCP_INFO is called TCP_CONNECTION_INFO, so for convenience these ones were defined to point to their equivalent. However there is a small difference now in that tcpi_rtt is called tcpi_rttcur on this platform, which forces us to make a special case for it before other platforms.	2022-04-15 17:51:09 +02:00
David CARLIER	5c83e3a156	MINOR: tcp_sample: clarifying samples support per os, for further expansion. While there is some overlap between what each OS provides in terms of retrievable info, each set is not a real subset of another one and this results in increasing complexity when trying to add support for new OSes. Let's just condition each item to the OS that support it. It's not pretty but at least it will avoid a real mess later. Note that fc_rtt and fc_rttvar are supported on any OS that has TCP_INFO, not just linux/freebsd/netbsd, so we continue to expose them unconditionally.	2022-04-15 17:51:09 +02:00
Christopher Faulet	39e436e222	BUG/MEDIUM: compression: Don't forget to update htx_sl and http_msg flags If the response is compressed, we must update the HTX start-line flags and the HTTP message flags. It is especially important if there is another filter enabled. Otherwise, there is no way to know the C-L header was removed and T-E one was added. Except by looping on headers. This patch is related to the issue #1660. It must backported as far as 2.0 (for HTX part only).	2022-04-15 16:22:33 +02:00
Christopher Faulet	32af9a7830	BUG/MEDIUM: fcgi-app: Use http_msg flags to know if C-L header can be added Instead of relying on the HTX start-line flags, it is better to rely on http_msg flags to know if a content-length header can be added or not. In addition, if the header is added, HTTP_MSGF_CNT_LEN flag must be added. Because of this bug, an invalid message can be emitted when the response is compressed because it may contain C-L and a T-E headers. This patch should fix the issue #1660. It must be backported as far as 2.2.	2022-04-15 16:11:55 +02:00
Amaury Denoyelle	f7ff9cbfe1	BUG/MEDIUM: quic: properly clean frames on stream free A released qc_stream_desc is freed as soon as all its buffer content has been acknowledged. However, it may still contains other frames waiting for ACK pointing to deleted buffer content. This can happen on retransmission. When freeing a qc_stream_desc, free all its frames in acked_frms tree to fix memory leak. This may also possibly fix a crash on retransmission. Now, the frames are properly removed from a packet. This ensure we do not retransmit a frame whose buffer is deallocated.	2022-04-15 13:45:28 +02:00
Christopher Faulet	2bb5edcf19	BUG/MEDIUM: connection: Don't crush context pointer location if it is a CS The issue only concerns the backend connection. The conn-stream is now owned by the stream and persists during all the stream life. Thus we must not crush it when the backend connection is released. It is 2.6-specific. No backport is needed.	2022-04-15 10:57:11 +02:00
Willy Tarreau	cef08c20c7	MINOR: extcheck: fill in the server's UNIX socket address when known While it's often a pain to try to figure a UNIX socket address, the server ones are reliable and may be emitted in the check provided they are retrieved in time. We cannot rely on addr_to_str() because it only reports "unix" since it may be used to log client addresses or listener addresses (which are renamed). The address length was extended to 256 chars to deal with long paths as previously it was limited to INET6_ADDRSTRLEN+1. This addresses github issue #101. There's no point backporting this, external checks are almost never used.	2022-04-14 19:56:32 +02:00
Willy Tarreau	c7edc9880a	CLEANUP: extcheck: do not needlessly preset the server's address/port During the config parsing we preset the server's address and port, but that's pointless since it's replaced during each check in order to deal with the possibility that the address was changed since.	2022-04-14 19:54:50 +02:00
Willy Tarreau	a544c66716	BUG/MEDIUM: stream: do not abort connection setup too early Github issue #472 reports a problem with short client connections making stick-table entries disappear. The problem is in fact totally different and stems at the connection establishment step. What happens is that the stick-table there has a single entry. The "stick-on" directive is forced to purge an existing entry before being able to create a new one. The new entry will be committed during the call to process_store_rules() on the response path. But if the client sends the FIN immediately after the connection is set up (e.g. using nc -z) then the SHUTR is received and will cancel the connection setup just after it starts. This cancellation will induce a call to cs_shutw() which will in turn leave the server-side state in ST_DIS. This transition from ST_CON to ST_DIS doesn't belong to the list of handled transition during the connection setup so it will be handled right after on the regular path, causing the connection to be closed. Because of this, we never pass through back_establish() and the backend's analysers are never set on the response channel, which is why process_store_rules() is not called and the stick-tables entry never committed. The comment above the code that causes this transition clearly says that the function is to be used after the connection is established with the server, but there's no such protection, and we always have the AUTO_CLOSE flag there (but there's hardly any available condition to eliminate it). This patch adds a test for the connection not being in ST_CON or for option abortonclose being set. It's sufficient to do the job and it should not cause issues. One concern was that the transition could happen during cs_recv() after the connection switches from CON to RDY then the read0 would be taken into account and would cause DIS to appear, which is not handled either. But that cannot happen because cs_recv() doesn't do anything until it's in ST_EST state, hence the read0() cannot be called from CON/RDY. Thus the transition from CON to DIS is only possible in back_handle_st_con() and back_handle_st_rdy() both of which are called when dealing with the transition already, or when abortonclose is set and the client aborts before connect() succeeds. It's possible that some further improvements could be made to detect this specific transition but it doesn't seem like anything would have to be added. This issue was first reported on 2.1. The abortonclose area is very sensitive so it would be wise to backport slowly, and probably no further than 2.4.	2022-04-14 17:39:48 +02:00
Amaury Denoyelle	5d774dee55	MINOR: quic: emit CONNECTION_CLOSE on app init error Emit a CONNECTION_CLOSE if the app layer cannot be properly initialized on qc_xprt_start. This force the quic-conn to enter the closing state before being closed. Without this, quic-conn normal operations continue, despite the app-layer reported as not initialized. This behavior is undefined, in particular when handling STREAM frames.	2022-04-14 15:09:32 +02:00
Amaury Denoyelle	05d4ae6436	BUG/MINOR: quic: fix return value for error in start Fix the return value used in quic-conn start callback for error. The caller expects a negative value in this case. Without this patch, the quic-conn and the connection stack are not closed despite an initialization failure error, which is an undefined behavior and may cause a crash in the end.	2022-04-14 15:08:16 +02:00
Amaury Denoyelle	622ec4166b	BUG/MINOR: quic-sock: do not double free session on conn init failure In the quic_session_accept, connection is in charge to call the quic-conn start callback. If this callback fails for whatever reason, there is a crash because of an explicit session_free. This happens because the connection is now the owner of the session due to previous conn_complete_session call. It will automatically calls session_free. Fix this by skipping the session_free explicit invocation on error. In practice, currently this has never happened as there is only limited cases of failures for conn_xprt_start for QUIC.	2022-04-14 14:50:12 +02:00
Amaury Denoyelle	2461bd534a	BUG/MINOR: mux-quic: prevent a crash in session_free on mux.destroy Implement qc_destroy. This callback is used to quickly release all MUX resources. session_free uses this callback. Currently, it can only be called if there was an error during connection initialization. If not defined, the process crashes.	2022-04-14 14:50:12 +02:00
Christopher Faulet	67df95a8a2	BUILD: http-client: Avoid dead code when compiled without SSL support When an HTTP client is started on an HAProxy compiled without the SSL support, an error is triggered when HTTPS is used. In this case, the freshly created conn-stream is released. But this code is specific to the non-SSL part. Thus it is moved the in right #if/#else section. This patch should fix the issue #1655.	2022-04-14 12:02:35 +02:00
Christopher Faulet	ae660be547	BUG/MEDIUM: mux-h1: Don't request more room on partial trailers The commit `744451c7c` ("BUG/MEDIUM: mux-h1: Properly detect full buffer cases during message parsing") introduced a regression if trailers are not received in one time. Indeed, in this case, nothing is appended in the channel buffer, while there are some data in the input buffer. In this case, we must not request more room to the upper layer, especially because the channel buffer can be empty. To fix the issue, on trailers parsing, we consider the H1 stream as congested when the max size allowed is reached. Of course, the H1 stream is also considered as congested if the trailers are too big and the channel buffer is not empty. This patch should fix the issue #1657. It must be backported as far as 2.0.	2022-04-14 11:57:06 +02:00
Christopher Faulet	cea05437c0	MINOR: conn-stream: Use unsafe functions to get conn/appctx in cs_detach_endp There is no reason to rely on safe functions here. This patch should fix the issue #1656.	2022-04-14 11:57:06 +02:00
Christopher Faulet	4de1bff866	MINOR: muxes: Don't expect to call release function with no mux defined For all muxes, the function responsible to release a mux is always called with a defined mux. Thus there is no reason to test if it is defined or not. Note the patch may seem huge but it is just because of indentation changes.	2022-04-14 11:57:06 +02:00
Christopher Faulet	4e61096e30	MINOR: muxes: Don't handle proto upgrade for muxes not supporting it Several muxes (h2, fcgi, quic) don't support the protocol upgrade. For these muxes, there is no reason to have code to support it. Thus in the destroy callback, there is now a BUG_ON() and the release function is simplified because the connection is always owned by the mux..	2022-04-14 11:57:06 +02:00
Christopher Faulet	7c452ccbff	MINOR: muxes: Don't expect to have a mux without connection in destroy callback Once a mux initialized, the underlying connection alwaus exists from its point of view and it is never removed until the mux is released. It may be owned by another mux during an upgrade. But the pointer remains set. Thus there is no reason to test it in the destroy callback function. This patch should fix the issue #1652.	2022-04-14 11:57:05 +02:00
Willy Tarreau	86b08a3e3e	BUG/MINOR: mux-h2: use timeout http-request as a fallback for http-keep-alive The doc states that timeout http-keep-alive is not set, timeout http-request is used instead. As implemented in commit `15a4733d5` ("BUG/MEDIUM: mux-h2: make use of http-request and keep-alive timeouts"), we use http-keep-alive unconditionally between requests, with a fallback on client/server. Let's make sure http-request is always used as a fallback for http-keep-alive first. This needs to be backported wherever the commit above is backported. Thanks to Christian Ruppert for spotting this.	2022-04-14 11:45:36 +02:00
Willy Tarreau	6ff91e2023	BUG/MINOR: mux-h2: do not use timeout http-keep-alive on backend side Commit `15a4733d5` ("BUG/MEDIUM: mux-h2: make use of http-request and keep-alive timeouts") omitted to check the side of the connection, and as a side effect, automatically enabled timeouts on idle backend connections, which is totally contrary to the principle that they must be autonomous. This needs to be backported wherever the patch above is backported.	2022-04-14 11:43:35 +02:00
Frédéric Lécaille	bc964bd1ae	BUG/MINOR: quic: Avoid starting the mux if no ALPN sent by the client If the client does not sent an ALPN, the SSL ALPN negotiation callback is not called. However, the handshake is reported as successful. Check just after SSL_do_handshake if an ALPN was negotiated. If not, emit a CONNECTION_CLOSE with a TLS alert to close the connection. This prevent a crash in qcc_install_app_ops() called with null as second parameter value.	2022-04-13 16:48:43 +02:00
Christopher Faulet	186354beac	MINOR: mux-h1: Rely on the endpoint instead of the conn-stream when possible Instead of testing if a conn-stream exists or not, we rely on CS_EP_ORPHAN endpoint flag. In addition, if possible, we access the endpoint from the h1s. Finally, the endpoint flags are now reported in trace messages.	2022-04-13 15:10:16 +02:00
Christopher Faulet	22050e0a2c	MINOR: muxes: Improve show_fd callbacks to dump endpoint flags H1, H2 and FCGI multiplexers define a show_fd callback to dump some internal info. The stream endpoint and its flags are now dumped if it exists.	2022-04-13 15:10:16 +02:00
Christopher Faulet	1336ccffab	CLEANUP: conn-stream: rename cs_register_applet() to cs_applet_create() cs_register_applet() was not a good name because it suggests it happens during startup, just like any other registration mechanisms..	2022-04-13 15:10:16 +02:00
Christopher Faulet	aa69d8fa1c	MINOR: conn-stream: Use a dedicated function to conditionally remove a CS cs_free_cond() must now be used to remove a CS. cs_free() may be used on error path to release a freshly allocated but unused CS. But in all other cases cs_free_cond() must be used. This function takes care to release the CS if it is possible (no app and detached from any endpoint). In fact, this function is only used internally. From the outside, cs_detach_* functions are used.	2022-04-13 15:10:16 +02:00
Christopher Faulet	a97ccedf6f	CLEANUP: muxes: Remove MX_FL_CLEAN_ABRT flag This flag is unused. Thus, it may be removed. No reason to still set it. It also cleans up "haproxy -vv" output.	2022-04-13 15:10:16 +02:00
Christopher Faulet	177a0e60ee	MEDIUM: check: Use a new conn-stream for each health-check run It is a partial revert of `54e85cbfc` ("MAJOR: check: Use a persistent conn-stream for health-checks"). But with the CS refactoring, the result is cleaner now. A CS is allocated when a new health-check run is started. The same CS is then used throughout the run. If there are several connections, the endpoint is just reset. At the end of the run, the CS is released. It means, in the tcp-check part, the CS is always defined.	2022-04-13 15:10:16 +02:00
Christopher Faulet	9ed7742673	DOC: conn-stream: Add comments on functions of the new CS api With the conn-stream refactoring, new functions were added. This patch adds missing comments to help devs to use them.	2022-04-13 15:10:16 +02:00
Christopher Faulet	265e165d82	CLEANUP: conn-stream: Don't export internal functions cs_new() and cs_attach_app() are only used internally. Thus, there is no reason to export them.	2022-04-13 15:10:16 +02:00
Christopher Faulet	6b0a0fb2f9	CLEANUP: tree-wide: Remove any ref to stream-interfaces Stream-interfaces are gone. Corresponding files can be safely be removed. In addition, comments are updated accordingly.	2022-04-13 15:10:16 +02:00
Christopher Faulet	582a226a2c	MINOR: conn-stream: Remove the stream-interface from the conn-stream The stream-interface API is no longer used. Thus, it is removed from the conn-stream. From now, stream-interfaces are now longer used !	2022-04-13 15:10:16 +02:00
Christopher Faulet	c77ceb6ad1	MEDIUM: stream: Don't use the stream-int anymore in process_stream() process_stream() and all associated functions now manipulate conn-streams. stream-interfaces are no longer used. In addition, function to dump info about a stream no longer print info about stream-interfaces.	2022-04-13 15:10:16 +02:00
Christopher Faulet	7739799ab4	MINOR: http-ana: Use CS to perform L7 retries do_l7_retry function now manipulated a conn-stream instead of a stream-interface.	2022-04-13 15:10:16 +02:00
Christopher Faulet	0eb32c0dd1	MINOR: stream: Use conn-stream to report server error the stream's srv_error callback function now manipulates a conn-stream instead of a stream-interface.	2022-04-13 15:10:16 +02:00
Christopher Faulet	5e29b76ea6	MEDIUM: stream-int/conn-stream: Move I/O functions to conn-stream cs_conn_io_cb(), cs_conn_sync_recv() and cs_conn_sync_send() are moved in conn_stream.c. Associated functions are moved too (cs_notify, cs_conn_read0, cs_conn_recv, cs_conn_send and cs_conn_process).	2022-04-13 15:10:15 +02:00
Christopher Faulet	a0bdec350f	MEDIUM: stream-int/conn-stream: Move blocking flags from SI to CS Remaining flags and associated functions are move in the conn-stream scope. These flags are added on the endpoint and not the conn-stream itself. This way it will be possible to get them from the mux or the applet. The functions to get or set these flags are renamed accordingly with the "cs_" prefix and updated to manipualte a conn-stream instead of a stream-interface.	2022-04-13 15:10:15 +02:00
Christopher Faulet	8f45eec016	MINOR: stream-int/conn-stream: Move si_alloc_ibuf() in the conn-stream scope si_alloc_ibuf() is renamed as c_alloc_ibuf() and update to manipulate a conn-stream instead of a stream-interface.	2022-04-13 15:10:15 +02:00
Christopher Faulet	158f33615d	MINOR: stream-int/conn-stream Move si_is_conn_error() in the conn-stream scope si_is_conn_error() is renamed as cs_is_conn_erro() and updated to manipulate a conn-stream instead of a stream-interface.	2022-04-13 15:10:15 +02:00
Christopher Faulet	000ba3e613	MINOR: conn-stream: Move si_conn_cb in the conn-stream scope si_conn_cb variable is renamed cs_data_conn_cb. In addtion, its associated functions are also renamed. si_cs_recv(), si_cs_send() and si_cs_process() are renamed cs_conn_recv(), cs_conn_send and cs_conn_process(). These functions are updated to manipulate conn-streams instead of stream-interfaces.	2022-04-13 15:10:15 +02:00
Christopher Faulet	431ce2e3c1	MINOR: stream-int/conn-stream: Move si_sync_recv/send() in conn-stream scope si_sync_recv() and si_sync_send() are renamesd cs_conn_recv() and cs_conn_send() and updated to manipulate conn-streams instead of stream-interfaces.	2022-04-13 15:10:15 +02:00
Christopher Faulet	4a7764ae9d	MINOR: stream-int/conn-stream: Move si_cs_io_cb() in the conn-stream scope si_cs_io_cb() is renamed cs_conn_io_cb(). In addition, the context of the tasklet used to wake-up the conn-stream is now a conn-stream.	2022-04-13 15:10:15 +02:00
Christopher Faulet	9029a72899	MINOR: stream-int/conn-stream: Move stream_int_notify() in the conn-stream scope stream_int_notify() is renamed cs_notify() and is updated to manipulate a conn-stream instead of a stream-interface.	2022-04-13 15:10:15 +02:00
Christopher Faulet	d715d36200	MINOR: stream-int/conn-stream: Move stream_int_read0() in the conn-stream scope stream_int_read0() is renamed cs_conn_read0() and is updated to manipulate a conn-stream instead of a stream-interface.	2022-04-13 15:10:15 +02:00
Christopher Faulet	6059ba4acc	MEDIUM: conn-stream/applet: Add a data callback for applets data callbacks were only used for streams attached to a connection and for health-checks. However there is a callback used by task_run_applet. So, si_applet_wake_cb() is first renamed to cs_applet_process() and it is defined as the data callback for streams attached to an applet. This way, this part now manipulates a conn-stream instead of a stream-interface. In addition, applets are no longer handled as an exception for this part.	2022-04-13 15:10:15 +02:00
Christopher Faulet	ef285c18f2	MINOR: stream-int/stream: Move si_update_both in stream scope si_update_both() is renamed stream_update_both_cs() and moved in stream.c. The function is slightly changed to manipulate the stream instead the front and back conn-streams.	2022-04-13 15:10:15 +02:00
Christopher Faulet	13045f0eae	MINOR: stream-int-conn-stream: Move si_update_* in conn-stream scope si_update_rx(), si_update_tx() and si_update() are renamed cs_update_rx(), cs_upate_tx() and cs_update() and updated to manipulate a conn-stream instead of a stream-interface.	2022-04-13 15:10:15 +02:00
Christopher Faulet	9ffddd5ca5	REORG: conn-stream: Move cs_app_ops in conn_stream.c Callback functions to perform shutdown for reads and writes and to trigger I/O calls are now moved in conn_stream.c.	2022-04-13 15:10:15 +02:00
Christopher Faulet	19bd728642	REORG: conn-stream: Move cs_shut* and cs_chk* in cs_utils cs_shutr(), cs_shutw(), cs_chk_rcv() and cs_chk_snd() are moved in cs_utils.h	2022-04-13 15:10:15 +02:00
Christopher Faulet	dde33046bd	REORG: stream-int: Move si_is_conn_error() in the header file To ease next changes, this function is moved in the header file. It is a transient commit.	2022-04-13 15:10:15 +02:00
Christopher Faulet	9b7a9b400d	REORG: stream-int: Export si_cs_recv(), si_cs_send() and si_cs_process() It is a transient commit. It should ease next changes about the conn-stream refactoring. At the end these functions will be moved in the conn-stream scope.	2022-04-13 15:10:15 +02:00
Christopher Faulet	aa91d6292b	MINOR: stream-int/connection: Move conn_si_send_proxy() in the connection scope conn_si_send_proxy() function is renamed conn_send_proxy() and moved in connection.c	2022-04-13 15:10:15 +02:00
Christopher Faulet	64b8d33577	MINOR: connection: unconst mux's get_fist_cs() callback function This change is mandatory for next commits.	2022-04-13 15:10:15 +02:00
Christopher Faulet	3704663e5f	MINOR: applet: Use the CS to register and release applets instead of SI si_register_applet() and si_applet_release() are renamed cs_register_applet() and cs_applet_release() and now manipulate a conn-stream instead of a stream-inteface.	2022-04-13 15:10:15 +02:00
Christopher Faulet	0c6a64cd5f	MEDIUM: stream-int/conn-stream: Move si_ops in the conn-stream scope The si_ops structure is renamed to cs_app_ops and the callback functions are changed to manipulate a conn-stream instead of a stream-interface..	2022-04-13 15:10:15 +02:00
Christopher Faulet	da098e6c17	MINOR: stream-int/conn-stream: Move si_shut* and si_chk* in conn-stream scope si_shutr(), si_shutw(), si_chk_rcv() and si_chk_snd() are moved in the conn-stream scope and renamed, respectively, cs_shutr(), cs_shutw(), cs_chk_rcv(), cs_chk_snd() and manipulate a conn-stream instead of a stream-interface.	2022-04-13 15:10:15 +02:00
Christopher Faulet	69ef6c9ef4	MINOR: conn-stream: Rename CS functions dedicated to connections Some conn-stream functions are only used when there is a connection. Thus, they was renamed with "cs_conn_" prefix. In addition, we expect to have a connection, so a BUG_ON is added to be sure the functions are never called in another context.	2022-04-13 15:10:15 +02:00
Christopher Faulet	2f35e7b6ab	MEDIUM: stream-int/conn-stream: Handle I/O subscriptions in the conn-stream wait_event structure is moved in the conn-stream. The tasklet is only created if the conn-stream is attached to a mux and released when the mux is detached. This implies a subtle change. In stream_int_chk_rcv() function, the wakeup of the tasklet was removed because there is no longer tasklet at this stage (stream_int_chk_rcv() is a callback function of si_embedded_ops).	2022-04-13 15:10:15 +02:00
Christopher Faulet	070b91bc11	MEDIUM: conn-stream: Be prepared to fail to attach a cs to a mux To be able to move wait_event from the stream-interface to the conn-stream, we must be prepare to handle errors when a mux is attached to a conn-stream. Indeed, the wait_event's tasklet will be allocated when both a mux and a stream will be both attached to a stream. So, we must be prepared to handle allocation errors.	2022-04-13 15:10:15 +02:00
Christopher Faulet	0797656ead	MINOR: conn-stream/connection: Move SHR/SHW modes in the connection scope These flags only concerns the connection part. In addition, it is required for a next commit, to avoid circular deps. Thus CS_SHR_* and CS_SHW_* were renamed with the "CO_" prefix.	2022-04-13 15:10:15 +02:00
Christopher Faulet	e39a4dfdf0	MINOR: stream-int/conn-stream: Move si_conn_ready() in the conn-stream scope si_conn_ready() is renamed cs_conn_ready() and handle a conn-stream insted of a stream-interface. The function is now in cs_utils.h.	2022-04-13 15:10:15 +02:00
Christopher Faulet	0a4dcb65ff	MINOR: stream-int/backend: Move si_connect() in the backend scope si_connect() is moved in backend.c and renamed as do_connect_server(). In addition, the function now manipulate a stream instead of a stream-interface.	2022-04-13 15:10:15 +02:00
Christopher Faulet	9125f3cc77	MINOR: stream-int/stream: Move si_retnclose() in the stream scope si_retnclose() is used to send a reply to a client before closing. There is no use on the server side, in spite of the function is generic. Thus, it is renamed stream_retnclose() and moved into the stream scope. The function now handle a stream and explicitly send a message to the client.	2022-04-13 15:10:15 +02:00
Christopher Faulet	62e757470a	MEDIUM: stream-int/conn-stream: Move stream-interface state in the conn-stream The stream-interface state (SI_ST_) is now in the conn-stream. It is a mechanical replacement for now. Nothing special. SI_ST_ and SI_SB_* were renamed accordingly. Utils functions to manipulate these infos were moved under the conn-stream scope. But it could be good to keep in mind that this part should be reworked. Indeed, at the CS level, we only need to know if it is ready to receive or to send. The state of conn-stream from INI to EST is only used on the server side. The client CS is immediately set to EST. Thus current SI_ST_* states should probably be moved to the stream to reflect the server connection state during the establishment stage.	2022-04-13 15:10:15 +02:00
Christopher Faulet	50264b41c8	MEDIUM: stream-int: Move SI err_type in the stream Only the server side is concerned by the stream-interface error type. It is useless to have an err_type field on the client side. So, it is now move to the stream. SI_ET_* are renames STRM_ET_* and moved in stream-t.h header file.	2022-04-13 15:10:14 +02:00
Christopher Faulet	a70a3548bc	MINOR: stream: Only save previous connection state for the server side The previous connection state on the client side was only used for debugging purpose to report client close. But this may be handled when the client stream-interface is switched from SI_ST_DIS to SI_ST_CLO. So, there only remains the previous connection state on the server side that is used by the stream, in process_stream(), to be able to set the correct termination flags. Thus, instead of keeping this info in the stream-interface for only one side, the info is now stored in the stream itself.	2022-04-13 15:10:14 +02:00
Christopher Faulet	78ed7f247b	CLEANUP: stream-int: Remove unused SI_FL_CLEAN_ABRT flag This flag is unused. So remove it to be able to remove the stream-interface.	2022-04-13 15:10:14 +02:00
Christopher Faulet	d139138bbc	MINOR: stream-int: Remove SI_FL_SRC_ADDR to rely on stream flags instead Flag to get the source ip/port with getsockname is now handled at the stream level. Thus SI_FL_SRC_ADDR stream-int flag is replaced by SF_SRC_ADDR stream flag.	2022-04-13 15:10:14 +02:00
Christopher Faulet	a728518c15	MINOR: stream-int: Remove SI_FL_INDEP_STR to rely on CS flags instead Flag to consider a stream as indepenent is now handled at the conn-stream level. Thus SI_FL_INDEP_STR stream-int flag is replaced by CS_FL_INDEP_STR conn-stream flags.	2022-04-13 15:10:14 +02:00
Christopher Faulet	974da9f8a4	MINOR: stream-int: Remove SI_FL_DONT_WAKE to rely on CS flags instead Flag to not wake the stream up on I/O is now handled at the conn-stream level. Thus SI_FL_DONT_WAKE stream-int flag is replaced by CS_FL_DONT_WAKE conn-stream flags.	2022-04-13 15:10:14 +02:00
Christopher Faulet	8abe712749	MINOR: stream-int: Remove SI_FL_NOLINGER/NOHALF to rely on CS flags instead Flags to disable lingering and half-close are now handled at the conn-stream level. Thus SI_FL_NOLINGER and SI_FL_NOHALF stream-int flags are replaced by CS_FL_NOLINGER and CS_FL_NOHALF conn-stream flags.	2022-04-13 15:10:14 +02:00
Christopher Faulet	ca2b5274b5	MINOR: mux-h2/mux-fcgi: Fully rely on CS_EP_KILL_CONN Instead of using a internal flag to kill the connection with the stream, we now fully rely on CS_EP_KILL_CONN flag.	2022-04-13 15:10:14 +02:00
Christopher Faulet	9a52123800	MINOR: stream-int: Remove SI_FL_KILL_CON to rely on conn-stream endpoint only Instead of setting a stream-interface flag to then set the corresponding conn-stream endpoint flag, we now only rely the conn-stream endoint. Thus SI_FL_KILL_CON is replaced by CS_EP_KILL_CONN. In addition si_must_kill_conn() is replaced by cs_must_kill_conn().	2022-04-13 15:10:14 +02:00
Christopher Faulet	7b5ca8f457	MINOR: channel: Use conn-streams as channel producer and consumer chn_prod() and chn_cons() now return a conn-stream instead of a stream-interface.	2022-04-13 15:10:14 +02:00
Christopher Faulet	6cd56d5a69	MEDIUM: conn-stream: Use endpoint error instead of conn-stream error Instead of relying on the conn-stream error, via CS_FL_ERR flags, we now directly use the error at the endpoint level with the flag CS_EP_ERROR. It should be safe to do so. But we must be careful because it is still possible that an error is processed too early. Anyway, a conn-stream has always a valid endpoint, maybe detached from any endpoint, but valid.	2022-04-13 15:10:14 +02:00
Christopher Faulet	af642df3b8	MINOR: stream-int/conn-stream: Report error to the CS instead of the SI SI_FL_ERR is removed and replaced by CS_FL_ERROR. It is a transient patch because the idea is to rely on the endpoint to handle errors at this level. But if for any reason it is not possible, the stream-interface flags will still be replaced.	2022-04-13 15:10:14 +02:00
Christopher Faulet	ae024ced03	MEDIUM: stream-int/stream: Use connect expiration instead of SI expiration The expiration date in the stream-interface was only used on the server side to set the connect, queue or turn-around timeout. It was checked on the frontend stream-interface, but never used concretely. So it was removed and replaced by a connect expiration date in the stream itself. Thus, SI_FL_EXP flag in stream-interfaces is replaced by a stream flag, SF_CONN_EXP.	2022-04-13 15:10:14 +02:00
Christopher Faulet	1d9877700e	MINOR: stream-int/conn-stream: Move half-close timeout in the conn-stream The half-close timeout (hcto) is now part of the conn-stream. It is a step closer to the stream-interface removal.	2022-04-13 15:10:14 +02:00
Christopher Faulet	8da67aae3e	MEDIUM: stream-int/conn-stream: Move src/dst addresses in the conn-stream The source and destination addresses at the applicative layer are moved from the stream-interface to the conn-stream. This simplifies a bit the code and it is a logicial step to remove the stream-interface.	2022-04-13 15:10:14 +02:00
Christopher Faulet	731c8e6cf9	MINOR: stream: Simplify retries counter calculation The conn_retries counter was set to the max value and decremented at each connection retry. Thus the counter reflected the number of retries left and not the real number of retries. All calculations of redispatch or reporting of number of retries experienced were made using subtracts from the configured retries, which was complicated and didn't bring any benefit. Now, this counter is set to 0 and incremented at each retry. We know we've reached the maximum allowed connection retries by comparing it to the configured value. In all other cases, we directly use the counter. This patch should address the feature request #1608.	2022-04-13 15:10:14 +02:00
Christopher Faulet	909f318259	MINOR: stream-int/stream: Move conn_retries counter in the stream The conn_retries counter may be moved into the stream structure. It only concerns the connection establishment. The frontend stream-interface does not use it. So it is a logical change.	2022-04-13 15:10:14 +02:00
Christopher Faulet	c216185c71	CLEANUP: http-ana: Remove http_alloc_txn() function Since the 2.4, this function is no longer used. Thus we can remove it.	2022-04-13 15:10:14 +02:00
Christopher Faulet	e05bf9e413	MINOR: stream-int/txn: Move buffer for L7 retries in the HTTP transaction The L7 retries only concerns the stream when a server connection is established. Thus instead of storing the L7 buffer into the stream-interface, it may be moved to the stream. And because it is only available for HTTP streams, it may be moved in the HTTP transaction. Associated flags are also moved into the HTTP transaction.	2022-04-13 15:10:14 +02:00
Christopher Faulet	908628c4c0	MEDIUM: tree-wide: Use CS util functions instead of SI ones At many places, we now use the new CS functions to get a stream or a channel from a conn-stream instead of using the stream-interface API. It is the first step to reduce the scope of the stream-interfaces. The main change here is about the applet I/O callback functions. Before the refactoring, the stream-interface was the appctx owner. Thus, it was heavily used. Now, as far as possible,the conn-stream is used. Of course, it remains many calls to the stream-interface API.	2022-04-13 15:10:14 +02:00
Christopher Faulet	3099511571	MINOR: conn-stream: Add ISBACK conn-stream flag CS_FL_ISBACK is a new flag, set on backend conn-streams. We must just be careful to preserve this flag when the endpoint is detached from the conn-stream.	2022-04-13 15:10:14 +02:00
Christopher Faulet	1bceee21e3	MINOR: mux-pt: Rely on the endpoint instead of the conn-stream when possible Instead of testing if a conn-stream exists or not, we rely on CS_EP_ORPHAN endpoint flag. In addition, if possible, we access the endpoint from the mux_pt context. Finally, the endpoint flags are now reported in trace messages.	2022-04-13 15:10:14 +02:00
Christopher Faulet	b041b23ae4	MEDIUM: conn-stream: Move remaning flags from CS to endpoint All old flags CS_FL_* are now moved in the endpoint scope and renamed CS_EP_* accordingly. It is a systematic replacement. There is no true change except for the health-check and the endpoint reset. Here it is a bit special because the same conn-stream is reused. Thus, we must handle endpoint allocation errors. To do so, cs_reset_endp() has been adapted. Thanks to this last change, it will now be possible to simplify the multiplexer and probably the applets too. A review must also be performed to remove some flags in the channel or the stream-interface. The HTX will probably be simplified too. Finally, there is now some place in the conn-stream to move info from the stream-interface.	2022-04-13 15:10:14 +02:00
Christopher Faulet	9ec2f4dc7c	MAJOR: conn-stream: Share endpoint struct between the CS and the mux/applet The conn-stream endpoint is now shared between the conn-stream and the applet or the multiplexer. If the mux or the applet is created first, it is responsible to also create the endpoint and share it with the conn-stream. If the conn-stream is created first, it is the opposite. When the endpoint is only owned by an applet or a mux, it is called an orphan endpoint (there is no conn-stream). When it is only owned by a conn-stream, it is called a detached endpoint (there is no mux/applet). The last entity that owns an endpoint is responsible to release it. When a mux or an applet is detached from a conn-stream, the conn-stream relinquishes the endpoint to recreate a new one. This way, the endpoint state is never lost for the mux or the applet.	2022-04-13 15:10:14 +02:00
Christopher Faulet	cb2fa368e9	REORG: applet: Uninline appctx_new function appctx_new() is moved in the C file and appctx_init() is now private.	2022-04-13 15:10:14 +02:00
Christopher Faulet	a9e8b3979d	MEDIUM: conn-stream: Pre-allocate endpoint to create CS from muxes and applets It is a transient commit to prepare next changes. Now, when a conn-stream is created from an applet or a multiplexer, an endpoint is always provided. In addition, the API to create a conn-stream was specialized to have one function per type. The next step will be to share the endpoint structure.	2022-04-13 15:10:14 +02:00
Christopher Faulet	b669d684c0	MEDIUM: conn-stream: Be able to pass endpoint to create a conn-stream It is a transient commit to prepare next changes. It is possible to pass a pre-allocated endpoint to create a new conn-stream. If it is NULL, a new endpoint is created, otherwise the existing one is used. There no more change at the conn-stream level. In the applets, all conn-stream are created with no pre-allocated endpoint. But for multiplexers, an endpoint is systematically created before creating the conn-stream.	2022-04-13 15:10:14 +02:00
Christopher Faulet	e9e4820288	MINOR: conn-stream: Move some CS flags to the endpoint Some CS flags, only related to the endpoint, are moved into the endpoint struct. More will probably moved later. Those ones are not critical. So it is pretty safe to move them now and this will ease next changes.	2022-04-13 15:10:14 +02:00
Christopher Faulet	db90f2aa9f	MEDIUM: conn-stream: Add an endpoint structure in the conn-stream Group the endpoint target of a conn-stream, its context and the associated flags in a dedicated structure in the conn-stream. It is not inlined in the conn-stream structure. There is a dedicated pool. For now, there is no complexity. It is just an indirection to get the endpoint or its context. But the purpose of this structure is to be able to share a refcounted context between the mux and the conn-stream. This way, it will be possible to preserve it when the mux is detached from the conn-stream.	2022-04-13 15:10:14 +02:00
Christopher Faulet	bb772d09f5	REORG: Initialize the conn-stream by hand in cs_init() The function cs_init() is only called by cs_new(). The conn-stream initialization will be reviewed. It is easier to do it in cs_new() instead of using a dedicated function. cs_new() is pretty simple, there is no reason to split the code in this case.	2022-04-13 15:10:14 +02:00
Christopher Faulet	9388204db1	MAJOR: conn-stream: Invert conn-stream endpoint and its context This change is only significant for the multiplexer part. For the applets, the context and the endpoint are the same. Thus, there is no much change. For the multiplexer part, the connection was used to set the conn-stream endpoint and the mux's stream was the context. But it is a bit strange because once a mux is installed, it takes over the connection. In a wonderful world, the connection should be totally hidden behind the mux. The stream-interface and, in a lesser extent, the stream, still access the connection because that was inherited from the pre-multiplexer era. Now, the conn-stream endpoint is the mux's stream (an opaque entity for the conn-stream) and the connection is the context. Dedicated functions have been added to attached an applet or a mux to a conn-stream.	2022-04-13 15:10:14 +02:00
Christopher Faulet	2479e5f775	MEDIUM: applet: Set the appctx owner during allocation The appctx owner is now always a conn-stream. Thus, it can be set during the appctx allocation. But, to do so, the conn-stream must be created first. It is not a problem on the server side because the conn-stream is created with the stream. On the client side, we must take care to create the conn-stream first. This change should ease other changes about the applets bootstrapping.	2022-04-13 15:10:13 +02:00
Christopher Faulet	81a40f630e	MINOR: conn-stream: Add flags to set the type of the endpoint This patch is mandatory to invert the endpoint and the context in the conn-stream. There is no common type (at least for now) for the entity representing a mux (h1s, h2s...), thus we must set its type when the endpoint is attached to a conn-stream. There is 2 types for the conn-stream endpoints: the mux (CS_FL_ENDP_MUX) and the applet (CS_FL_ENDP_APP).	2022-04-13 15:10:13 +02:00
Christopher Faulet	4aa1d2838c	MINOR: applet: Make .init callback more generic For now there is no much change. Only the appctx is passed as argument when the .init callback function is called. And it is not possible to yield at this stage. It is not a problem because the feature is not used. Only the lua defines this callback function for the lua TCP/HTTP services. The idea is to be able to use it for all applets to initialize the appctx context.	2022-04-13 15:10:13 +02:00
Christopher Faulet	91449b0351	BUG/MINOR: mux-h1: Don't release unallocated CS on error path cs_free() must not be called when we fail to allocate the conn-stream in h1s_new_cs() function. This bug was introduced by the commit `cda94accb` ("MAJOR: stream/conn_stream: Move the stream-interface into the conn-stream"). It is 2.6-specific, no backport is needed.	2022-04-13 15:10:13 +02:00
Willy Tarreau	f1de1b51ca	BUG/MINOR: cache: do not display expired entries in "show cache" It was mentioned in issue #12 that expired entries would appear with a negative expire delay in "show cache". Instead of listing them, let's just evict them. This could be backported to all versions since this was reported on 1.8 already.	2022-04-13 11:21:39 +02:00
Willy Tarreau	15dbedd63d	BUG/MINOR: mux-h2: do not send GOAWAY if SETTINGS were not sent It was reported in issue #13 that a GOAWAY frame was sent on timeout even if no SETTINGS frame was sent. The approach imagined by then was to track the fact that a SETTINGS frame was already sent to avoid this, but that's already what is done through the state, though it doesn't stand due to the fact that we switch the frame to the error state. Thus instead what we're doing here is to instead set the GOAWAY_FAILED flag in h2c_error() before switching to the ERROR state when the state indicates we've not yet sent settings, and refrain from sending anything from the h2c_send_goaway_error() function for such states. This could be backported to all versions where it applies well.	2022-04-13 09:40:52 +02:00
Amaury Denoyelle	bb97042254	BUG/MINOR: h3: fix build with DEBUG_H3 qcs by_id field has been replaced by a new field named "id". Adjust the h3_debug_printf traces. This is the case since the introduction of the qc_stream_desc type.	2022-04-12 16:42:45 +02:00
Willy Tarreau	3b75748542	BUILD/DEBUG: hpack: use unsigned int in printf format in debug code In issue #1184 cppcheck found that the debug code incorrectly uses %d to print an unsigned value.	2022-04-12 08:40:38 +02:00
Willy Tarreau	160e74bb9e	BUILD/DEBUG: hpack-tbl: fix format string in standalone debug code In issue #1184, cppcheck reports that an incorrect format "%d" was used to print an unsigned in the debug code, though values are always very small and this will never be an issue.	2022-04-12 08:30:08 +02:00
Willy Tarreau	2645b34341	BUILD: peers: adjust some printf format to silence cppcheck In issue #1184, cppcheck complains about some inconsistent printf formats. At least the one in peer_prepare_hellomsg() that uses "%u" for the int "min_ver" is wrong. Let's force other types to make it happy, though constants cannot cause trouble.	2022-04-12 08:28:18 +02:00
Willy Tarreau	f0683cd510	BUILD/DEBUG: lru: fix printf format in debug code cppcheck reports in issue #1184 a type mismatch between "%d" and the unsigned int "misses" in the standalone debug code of lru.c. Let's switch to "%u".	2022-04-12 08:19:33 +02:00
Willy Tarreau	807a3a53bb	MINOR: log: add '~' to frontend when the transport layer provides SSL We used to check if the transport layer was ssl_sock to decide to log "~" after a frontend's name. Now that QUIC is present, this doesn't work anymore. Better rely on the transport layer's get_ssl_sock_ctx() method.	2022-04-12 08:08:33 +02:00
Willy Tarreau	8f6c0c32b8	BUG/MINOR: sock: do not double-close the accepted socket on the error path Coverity found in issue #1646 that I added a double-close bug in last commit `e4d09cedb` ("MINOR: sock: check configured limits at the sock layer, not the listener's") because the error path already closes the FD. No backport needed.	2022-04-12 07:49:11 +02:00
Willy Tarreau	ce7a5e0967	MINOR: ssl: refine the error testing for fc_err and fc_err_str In issue #1645, coverity suspects some dead code due to a pair of remaining tests on "if (!ctx)". While all other functions test the context earlier, these ones used to only test the connection and the transport. It's still not very clear to me if there are certain error cases that can lead to no SSL being initially set while the rest is ready, and the SSL arriving later, but better preserve this original construct by testing first the connection and only later the context.	2022-04-12 07:42:49 +02:00
Willy Tarreau	3a0a0d6cc1	BUILD: ssl: add an unchecked version of __conn_get_ssl_sock_ctx() First gcc, then now coverity report possible null derefs in situations where we know these cannot happen since we call the functions in contexts that guarantee the existence of the connection and the method used. Let's introduce an unchecked version of the function for such cases, just like we had to do with objt_*. This allows us to remove the ALREADY_CHECKED() statements (which coverity doesn't see), and addresses github issues #1643, #1644, #1647.	2022-04-12 07:33:26 +02:00
Willy Tarreau	99ade09cbf	BUILD: ssl: fix build warning with previous changes to ssl_sock_ctx Some compilers see a possible null deref after conn_get_ssl_sock_ctx() in ssl_sock_parse_heartbeat, which cannot happen there, so let's mark it as safe. No backport needed.	2022-04-11 19:47:31 +02:00
Willy Tarreau	784b868c97	MEDIUM: quic: move conn->qc into conn->handle It was supposed to be there, and probably was not placed there due to historic limitations in listener_accept(), but now there does not seem to be a remaining valid reason for keeping the quic_conn out of the handle. In addition in new_quic_cli_conn() the handle->fd was incorrectly set to the listener's FD.	2022-04-11 19:33:04 +02:00
Willy Tarreau	54a1dcb1bb	MEDIUM: xprt-quic: implement get_ssl_sock_ctx() By being able to return the ssl_sock_ctx, we're now enabling the whole set of SSL sample fetch methods to work on the current SSL context of the QUIC connection, as seen in the following test showing a request forwarded to an HTTP/1 server with plenty of SSL headers filled: 00000001:decrypt.clireq[000f:ffffffff]: GET / HTTP/1.1 00000001:decrypt.clihdr[000f:ffffffff]: host: localhost 00000001:decrypt.clihdr[000f:ffffffff]: user-agent: nghttp3/ngtcp2 client 00000001:decrypt.clihdr[000f:ffffffff]: x-src: 127.0.0.1 00000001:decrypt.clihdr[000f:ffffffff]: x-dst: 127.0.0.4 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_f_serial: D16197E7D3E634E9 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_f_key_alg: rsaEncryption 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_f_sig_alg: RSA-SHA1 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_fc: 1 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_fc_has_sni: 1 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_fc_sni: blah 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_fc_alpn: h3 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_fc_protocol: TLSv1.3 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_fc_cipher: TLS_AES_256_GCM_SHA384 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_fc_alg_keysize: 256 00000001:decrypt.clihdr[000f:ffffffff]: x-ssl_fc_use_keysize: 256 00000001:decrypt.clihdr[000f:ffffffff]: x-forwarded-for: 127.0.0.1 The code is trivial, but this is marked as medium as there's always the risk that some of the callable functions do not like being called on such SSL contexts.	2022-04-11 19:33:04 +02:00
Willy Tarreau	939b0bf866	MEDIUM: ssl: stop using conn->xprt_ctx to access the ssl_sock_ctx The SSL functions must not use conn->xprt_ctx anymore but find the context by calling conn_get_ssl_sock_ctx(), which will properly pass through the transport layers to retrieve the desired information. Otherwise when the functions are called on a QUIC connection, they refuse to work for not being called on the proper transport.	2022-04-11 19:33:04 +02:00
Willy Tarreau	de827958a2	MEDIUM: ssl: improve retrieval of ssl_sock_ctx and SSL detection Historically there was a single way to have an SSL transport on a connection, so detecting if the transport layer was SSL and a context was present was sufficient to detect SSL. With QUIC, things have changed because QUIC also relies on SSL, but the context is embedded inside the quic_conn and the transport layer doesn't match expectations outside, making it difficult to detect that SSL is in use over the connection. The approach taken here to improve this consists in adding a new method at the transport layer, get_ssl_sock_ctx(), to retrieve this often needed ssl_sock_ctx, and to use this to detect the presence of SSL. This will even allow some simplifications and cleanups to be made in the SSL code itself, and QUIC will be able to provide one to export its ssl_sock_ctx.	2022-04-11 19:33:04 +02:00
Willy Tarreau	cdf7c8e543	MINOR: quic-sock: provide a pair of get_src/get_dst functions These functions will allow the connection layer to retrieve a quic_conn's source or destination when possible. The quic_conn holds the peer's address but not the local one, and the sockets API doesn't always makes that easy for datagrams. Thus for frontend connection what we're doing here is to retrieve the listener's address when the destination address is desired. Now it finally becomes possible to fetch the source and destination using "src" and "dst", and to pass an incoming connection's endpoints via the proxy protocol.	2022-04-11 19:33:04 +02:00
Willy Tarreau	671bd5af25	MINOR: mux-quic: properly set the flags and name fields The mux didn't have its flags nor name set, as seen in this output of "haproxy -vv": Available multiplexer protocols : (protocols marked as <default> cannot be specified using 'proto' keyword) quic : mode=HTTP side=FE mux= flags= h2 : mode=HTTP side=FE\|BE mux=H2 flags=HTX\|CLEAN_ABRT\|HOL_RISK\|NO_UPG This might have random impacts at certain points like forcing some connections to close instead of aborting a stream, or not always handling certain streams as fully HTX-compliant.	2022-04-11 19:32:51 +02:00
Willy Tarreau	07ecfc5e88	MEDIUM: connection: panic when calling FD-specific functions on FD-less conns Certain functions cannot be called on an FD-less conn because they are normally called as part of the protocol-specific setup/teardown sequence. Better place a few BUG_ON() to make sure none of them is called in other situations. If any of them would trigger in ambiguous conditions, it would always be possible to replace it with an error.	2022-04-11 19:31:47 +02:00
Willy Tarreau	e22267971b	MINOR: connection: skip FD-based syscalls for FD-less connections Some syscalls at the TCP level act directly on the FD. Some of them are used by TCP actions like set-tos, set-mark, silent-drop, others try to retrieve TCP info, get the source or destination address. These ones must not be called with an invalid FD coming from an FD-less connection, so let's add the relevant tests for this. It's worth noting that all these ones already have fall back plans (do nothing, error, or switch to alternate implementation).	2022-04-11 19:31:47 +02:00
Willy Tarreau	0e9c264ca0	MINOR: connection: use conn_fd() when displaying connection errors The SSL connection errors and socks4 proxy errors used to blindly dump the FD, now it's sanitized via conn_fd().	2022-04-11 19:31:47 +02:00
Willy Tarreau	a57f34523e	MINOR: stream: only dump connections' FDs when they are valid The "show sess" output and the debugger outputs will now use conn_fd() to retrieve the file descriptor instead of dumping incorrect data.	2022-04-11 19:31:47 +02:00
Willy Tarreau	c78a9698ef	MINOR: connection: add a new flag CO_FL_FDLESS on fd-less connections QUIC connections do not use a file descriptor, instead they use the quic equivalent which is the quic_conn. A number of our historical functions at the connection level continue to unconditionally touch the file descriptor and this may have consequences once QUIC starts to be used. This patch adds a new flag on QUIC connections, CO_FL_FDLESS, to mention that the connection doesn't have a file descriptor, hence the FD-based API must never be used on them. From now on it will be possible to intrument existing functions to panic when this flag is present.	2022-04-11 19:31:47 +02:00
Willy Tarreau	e4d09cedb6	MINOR: sock: check configured limits at the sock layer, not the listener's listener_accept() used to continue to enforce the FD limits relative to global.maxsock by itself while it's the last FD-specific test in the whole file. This test has nothing to do there, it ought to be placed in sock_accept_conn() which is the one in charge of FD allocation and tests. Similar tests are already located there by the way. The only tiny difference is that listener_accept() used to pause for one second when this limit was reached, while other similar conditions were pausing only 100ms, so now the same 100ms will apply. But that's not important and could even be considered as an improvement.	2022-04-11 19:31:47 +02:00
Willy Tarreau	325fc63f5a	BUILD: xprt-quic: replace ERR_func_error_string() with ERR_peek_error_func() OpenSSL 3.0 warns that ERR_func_error_string() is deprecated. Using ERR_peek_error_func() solves it instead, and this function was added to the compat layer by commit `1effd9aa0` ("MINOR: ssl: Remove call to ERR_func_error_string with OpenSSLv3").	2022-04-11 18:54:46 +02:00
William Lallemand	d7bfbe2333	BUILD: ssl: add USE_ENGINE and disable the openssl engine by default The OpenSSL engine API is deprecated starting with OpenSSL 3.0. In order to have a clean build this feature is now disabled by default. It can be reactivated with USE_ENGINE=1 on the build line.	2022-04-11 18:41:24 +02:00
Willy Tarreau	00147f7244	BUG/MINOR: stats: define the description' background color in dark color scheme Shawn Heisey reported that the proxy's description was unreadable in dark color scheme. This is because the text color is changed in the table but not the cell's background. This should be backported to 2.5.	2022-04-11 08:01:43 +02:00
Willy Tarreau	dddfec5428	CLEANUP: connection: reduce the with of the mux dump output In "haproxy -vv" we produce a list of available muxes with their capabilities, but that list is often quite large for terminals due to excess of spaces, so let's reduce them a bit to make the output more readable.	2022-04-09 11:43:16 +02:00
Remi Tricot-Le Breton	b5d968d9b2	MEDIUM: global: Add a "close-spread-time" option to spread soft-stop on time window The new 'close-spread-time' global option can be used to spread idle and active HTTP connction closing after a SIGUSR1 signal is received. This allows to limit bursts of reconnections when too many idle connections are closed at once. Indeed, without this new mechanism, in case of soft-stop, all the idle connections would be closed at once (after the grace period is over), and all active HTTP connections would be closed by appending a "Connection: close" header to the next response that goes over it (or via a GOAWAY frame in case of HTTP2). This patch adds the support of this new option for HTTP as well as HTTP2 connections. It works differently on active and idle connections. On active connections, instead of sending systematically the GOAWAY frame or adding the 'Connection: close' header like before once the soft-stop has started, a random based on the remainder of the close window is calculated, and depending on its result we could decide to keep the connection alive. The random will be recalculated for any subsequent request/response on this connection so the GOAWAY will still end up being sent, but we might wait a few more round trips. This will ensure that goaways are distributed along a longer time window than before. On idle connections, a random factor is used when determining the expire field of the connection's task, which should naturally spread connection closings on the time window (see h2c_update_timeout). This feature request was described in GitHub issue #1614. This patch should be backported to 2.5. It depends on "BUG/MEDIUM: mux-h2: make use of http-request and keep-alive timeouts" which refactorized the timeout management of HTTP2 connections.	2022-04-08 18:15:21 +02:00
Frédéric Lécaille	8c7927c6dd	MINOR: quic_tls: Make key update use of reusable cipher contexts We modify the key update feature implementation to support reusable cipher contexts as this is done for the other cipher contexts for packet decryption and encryption. To do so we attach a context to the quic_tls_kp struct and initialize it each time the underlying secret key is updated. Same thing when we rotate the secrets keys, we rotate the contexts as the same time.	2022-04-08 15:38:29 +02:00
Frédéric Lécaille	3dfd4c4b0d	MINOR: quic: Add short packet key phase bit values to traces This is useful to diagnose key update related issues.	2022-04-08 15:38:29 +02:00
Frédéric Lécaille	9688a8df49	CLEANUP: quic: Do not set any cipher/group from ssl_quic_initial_ctx() These settings are potentially cancelled by others setting initialization shared with SSL sock bindings. This will have to be clarified when we will adapt the QUIC bindings configuration.	2022-04-08 15:38:29 +02:00
Frédéric Lécaille	f2f4a4eee5	MINOR: quic_tls: Stop hardcoding cipher IV lengths For QUIC AEAD usage, the number of bytes for the IVs is always 12.	2022-04-08 15:38:29 +02:00
Frédéric Lécaille	f4605748f4	MINOR: quic_tls: Add reusable cipher contexts to QUIC TLS contexts Add ->ctx new member field to quic_tls_secrets struct to store the cipher context for each QUIC TLS context TX/RX parts. Add quic_tls_rx_ctx_init() and quic_tls_tx_ctx_init() functions to initialize these cipher context for RX and TX parts respectively. Make qc_new_isecs() call these two functions to initialize the cipher contexts of the Initial secrets. Same thing for ha_quic_set_encryption_secrets() to initialize the cipher contexts of the subsequent derived secrets (ORTT, Handshake, 1RTT). Modify quic_tls_decrypt() and quic_tls_encrypt() to always use the same cipher context without allocating it each time they are called.	2022-04-08 15:38:29 +02:00
Frédéric Lécaille	82851bd3cb	BUG/MEDIUM: quic: Possible crash from quic_free_arngs() All quic_arng_node objects are allocated from "pool_head_quic_arng" memory pool. They must be deallocated calling pool_free().	2022-04-08 15:38:29 +02:00
Willy Tarreau	9cc88c3075	BUG/MINOR: quic: set the source not the destination address on accept() When a QUIC connection is accepted, the wrong field is set from the client's source address, it's the destination instead of the source. No backport needed.	2022-04-08 14:43:27 +02:00
Amaury Denoyelle	8038821c88	BUG/MEDIUM: mux-quic: properly release conn-stream on detach On qc_detach(), the qcs must cleared the conn-stream context and set its cs pointer to NULL. This prevents the qcs to point to a dangling reference. Without this, a SEGFAULT may occurs in qc_wake_some_streams() when accessing an already detached conn-stream instance through a qcs. Here is the SEGFAULT observed on haproxy.org. Program terminated with signal 11, Segmentation fault. 1234 else if (qcs->cs->data_cb->wake) { (gdb) p qcs.cs.data_cb $1 = (const struct data_cb *) 0x0 This can happens since the following patch : commit `fe035eca3a` MEDIUM: mux-quic: report errors on conn-streams	2022-04-08 12:09:23 +02:00
Amaury Denoyelle	9c3955c98c	CLEANUP: mux-quic: remove uneeded TODO in qc_detach The stream mux buffering has been reworked since the introduction of the struct qc_stream_desc. A qcs is now able to quickly release its buffer to the quic-conn.	2022-04-08 12:08:50 +02:00
Christopher Faulet	114e759d5d	BUG/MEDIUM: http-act: Don't replace URI if path is not found or invalid For replace-path, replace-pathq and replace-uri actions, we must take care to not match on the selected element if it is not defined. regex_exec_match2() function expects to be called with a defined subject. However, if the request path is invalid or not found, the function is called with a NULL subject, leading to a crash when compiled without the PRCE/PCRE2 support. For instance the following rules crashes HAProxy on a CONNECT request: http-request replace-path /short/(.) /\1 This patch must be backported as far as 2.0.	2022-04-08 10:45:31 +02:00
Christopher Faulet	21ac0eec28	BUG/MEDIUM: http-conv: Fix url_enc() to not crush const samples url_enc() encodes an input string by calling encode_string(). To do so, it adds a trailing '\0' to the sample string. However it never restores the sample string at the end. It is a problem for const samples. The sample string may be in the middle of a buffer. For instance, the HTTP headers values are concerned. However, instead of modifying the sample string, it is easier to rely on encode_chunk() function. It does the same but on a buffer. This patch must be backported as far as 2.2.	2022-04-08 10:12:59 +02:00
Christopher Faulet	dca3b5b2c6	BUG/MINOR: http_client: Don't add input data on an empty request buffer When compiled in debug mode, a BUG_ON triggers an error when the payload is fully transfered from the http-client buffer to the request channel buffer. In fact, when channel_add_input() is called, the request buffer is empty. So an error is reported when those data are directly forwarded, because we try to add some output data on a buffer with no data. To fix the bug, we must be sure to call channel_add_input() after the data transfer. The bug was introduced by the commit `ccc7ee45f` ("MINOR: httpclient: enable request buffering"). So, this patch must be backported if the above commit is backported.	2022-04-07 11:04:39 +02:00
Christopher Faulet	2eb5243e7f	BUG/MEDIUM: mux-h1: Set outgoing message to DONE when payload length is reached When a message is sent, we can switch it state to MSG_DONE if all the announced payload was processed. This way, if the EOM flag is not set on the message when the last expected data block is processed, the message can still be set to MSG_DONE state. This bug is related to the previous ones. There is a design issue with the HTX since the 2.4. When the EOM HTX block was replaced by a flag, I tried hard to be sure the flag is always set with the last HTX block on a message. It works pretty well for all messages received from a client or a server. But for internal messages, it is not so simple. Because applets cannot always properly handle the end of messages. So, there are some cases where the EOM flag is set on an empty message. As a workaround, for chunked messages, we can add an EOT HTX block. It does the trick. But for messages with a content-length, there is no empty DATA block. Thus, the only way to be sure the end of the message was reached in this case is to detect it in the H1 multiplexr. We already count the amount of data processed when the payload length is announced. Thus, we must only switch the message in DONE state when last bytes of the payload are received. Or when the EOM flag is received of course. This patch must be backported as far as 2.4.	2022-04-07 11:04:07 +02:00
Christopher Faulet	f89af9c205	BUG/MEDIUM: hlua: Don't set EOM flag on an empty HTX message in HTTP applet In a lua HTTP applet, when the script is finished, we must be sure to not set the EOM on an empty message. Otherwise, because there is no data to send, the mux on the client side may miss the end of the message and consider any shutdown as an abort. See "UG/MEDIUM: stats: Be sure to never set EOM flag on an empty HTX message" for details. This patch must be backported as far as 2.4. On previous version there is still the EOM HTX block.	2022-04-07 11:04:07 +02:00
Christopher Faulet	08b45cb8fd	BUG/MEDIUM: stats: Be sure to never set EOM flag on an empty HTX message During the last call to the stats I/O handle, it is possible to have nothing to dump. It may happen for many reasons. For instance, all remaining proxies are disabled or they don't match the specified scope. In HTML or in JSON, it is not really an issue because there is a footer. So there are still some data to push in the response channel buffer. In CSV, it is a problem because there is no footer. It means it is possible to finish the response with no payload at all in the HTX message. Thus, the EOM flag may be added on an empty message. When this happens, a shutdown is performed on an empty HTX message. Because there is nothing to send, the mux on the client side is not notified that the message was properly finished and interprets the shutdown as an abort. The response is chunked. So an abort at this stage means the last CRLF is never sent to the client. All data were sent but the message is invalid because the response chunking is not finished. If the reponse is compressed, because of a similar bug in the comppression filter, the compression is also aborted and the content is truncated because some data a lost in the compression filter. It is design issue with the HTX. It must be addressed. And there is an opportunity to do so with the recent conn-stream refactoring. It must be carefully evaluated first. But it is possible. In the means time and to also fix stable versions, to workaround the bug, a end-of-trailer HTX block is systematically added at the end of the message when the EOM flag is set if the HTX message is empty. This way, there are always some data to send when the EOM flag is set. Note that with a H2 client, it is only a problem when the response is compressed. This patch should fix the issue #1478. It must be backported as far as 2.4. On previous versions there is still the EOM block.	2022-04-07 11:04:07 +02:00
Christopher Faulet	d057960769	BUG/MINOR: fcgi-app: Don't add C-L header on response to HEAD requests In the FCGI app, when a full response is received, if there is no content-length and transfer-encoding headers, a content-length header is automatically added. This avoid, as far as possible to chunk the response. This trick was added because, most of time, scripts don"t add those headers. But this should not be performed for response to HEAD requests. Indeed, in this case, there is no payload. If the payload size is not specified, we must not added it by hand. Otherwise, a "content-length: 0" will always be added while it is not the real payload size (unknown at this stage). This patch should solve issue #1639. It must be backported as far as 2.2.	2022-04-07 11:04:07 +02:00
Amaury Denoyelle	b515b0af1d	MEDIUM: quic: report closing state for the MUX Define a new API to notify the MUX from the quic-conn when the connection is about to be closed. This happens in the following cases : - on idle timeout - on CONNECTION_CLOSE emission or reception The MUX wake callback is called on these conditions. The quic-conn QUIC_FL_NOTIFY_CLOSE is set to only report once. On the MUX side, connection flags CO_FL_SOCK_RD_SH\|CO_FL_SOCK_WR_SH are set to interrupt future emission/reception. This patch is the counterpart to "MEDIUM: mux-quic: report CO_FL_ERROR on send". Now the quic-conn is able to report its closing, which may be translated by the MUX into a CO_FL_ERROR on the connection for the upper layer. This allows the MUX to properly react to the QUIC closing mechanism for both idle-timeout and closing/draining states.	2022-04-07 10:37:45 +02:00
Amaury Denoyelle	fe035eca3a	MEDIUM: mux-quic: report errors on conn-streams Complete the error reporting. For each attached streams, if CO_FL_ERROR is set, mark them with CS_FL_ERR_PENDING\|CS_FL_ERROR. This will notify the upper layer to trigger streams detach and release of the MUX. This reporting is implemented in a new function qc_wake_some_streams(), called by qc_wake(). This ensures that a lower-layer error is quickly reported to the individual streams.	2022-04-07 10:37:45 +02:00
Amaury Denoyelle	d97fc804f9	MEDIUM: mux-quic: report CO_FL_ERROR on send Mark the connection with CO_FL_ERROR on qc_send() if the socket Tx is closed. This flag is used by the upper layer to order a close on the MUX. This requires to check CO_FL_ERROR in qcc_is_dead() to process to immediate MUX free when set. The qc_wake() callback has been completed. Most notably, it now calls qc_send() to report a possible CO_FL_ERROR. This is useful because qc_wake() is called by the quic-conn on imminent closing. Note that for the moment the error flag can never be set because the quic-conn does not report when the Tx socket is closed. This will be implemented in a following patch.	2022-04-07 10:35:34 +02:00
Amaury Denoyelle	c933780f1e	MINOR: mux-quic: centralize send operations in qc_send Regroup all features related to sending in qc_send(). This will be useful when qc_send() will be called outside of the io-cb. Currently, flow-control frames generation is now automatically integrated in qc_send().	2022-04-07 10:23:10 +02:00
Amaury Denoyelle	198d35f9c6	MINOR: mux-quic: define is_active app-ops Add a new app layer operation is_active. This can be used by the MUX to check if the connection can be considered as active or not. This is used inside qcc_is_dead as a first check. For example on HTTP/3, if there is at least one bidir client stream opened the connection is active. This explicitly ignore the uni streams used for control and qpack as they can never be closed during the connection lifetime.	2022-04-07 10:23:10 +02:00
Amaury Denoyelle	06890aaa91	MINOR: mux-quic: adjust timeout to accelerate closing Improve timeout handling on the MUX. When releasing a stream, first check if the connection can be considered as dead and should be freed immediatly. This allows to liberate resources faster when possible. If the connection is still active, ensure there is no attached conn-stream before scheduling the timeout. To do this, add a nb_cs field in the qcc structure.	2022-04-07 10:23:10 +02:00
Amaury Denoyelle	846cc046ae	MINOR: mux-quic: factorize conn-stream attach Provide a new function qc_attach_cs. This must be used by the app layer when a conn-stream can be instantiated. This will simplify future development.	2022-04-07 10:10:23 +02:00
Amaury Denoyelle	c9acc31018	BUG/MINOR: fix memleak on quic-conn streams cleaning When freeing a quic-conn, the streams resources attached to it must be cleared. This code is already implemented but the streams buffer was not deallocated. Fix this by using the function qc_stream_desc_free. This existing function centralize all operations to properly free all streams elements, attached both to the MUX and the quic-conn. This fixes a memory leak which can happen for each released connection.	2022-04-07 10:10:23 +02:00
Amaury Denoyelle	6057b4090e	CLEANUP: mux-quic: remove unused QC_CF_CC_RECV This flag was used to notify the MUX about a CONNECTION_CLOSE frame reception. It is now unused on the MUX side and can be removed. A new mechanism to detect quic-conn closing will be soon implemented.	2022-04-07 10:10:23 +02:00
Amaury Denoyelle	e0be573c1b	CLEANUP: quic: use static qualifer on quic_close quic_close can be used through xprt-ops and can thus be kept as a static symbol.	2022-04-07 10:10:22 +02:00
Amaury Denoyelle	db71e3bd09	BUG/MEDIUM: quic: ensure quic-conn survives to the MUX Rationalize the lifetime of the quic-conn regarding with the MUX. The quic-conn must not be freed if the MUX is still allocated. This simplify the MUX code when accessing the quic-conn and removed possible segfaults. To implement this, if the quic-conn timer expired, the quic-conn is released only if the MUX is not allocated. Else, the quic-conn is flagged with QUIC_FL_CONN_EXP_TIMER. The MUX is then responsible to call quic_close() which will free the flagged quic-conn.	2022-04-07 10:10:22 +02:00
Frédéric Lécaille	59bf255806	MINOR: quic: Add closing connection state New received packets after sending CONNECTION_CLOSE frame trigger a new CONNECTION_CLOSE frame to be sent. Each time such a frame is sent we increase the number of packet required to send another CONNECTION_CLOSE frame. Rearm only one time the idle timer when sending a CONNECTION_CLOSE frame.	2022-04-06 15:52:35 +02:00
Frédéric Lécaille	47756809fb	MINOR: quic: Add draining connection state. As soon as we receive a CONNECTION_CLOSE frame, we must stop sending packets. We add QUIC_FL_CONN_DRAINING connection flag to do so.	2022-04-06 15:52:35 +02:00
William Lallemand	eb0d4c40ac	BUG/MINOR: httpclient: end callback in applet release In case an error provokes the release of the applet, we will never call the end callback of the httpclient. In the case of a lua script, it would mean that the lua task will never be waked up after a yield, letting the lua script stuck forever. Fix the issue by moving the callback from the end of the iohandler to the applet release function. Must be backported in 2.5.	2022-04-06 14:19:36 +02:00
William Lallemand	71abad050a	MEDIUM: httpclient: enable l7-retry Enable the layer-7 retry in the httpclient. This way the client will retry upon a connection error or a timeout.	2022-04-06 11:50:10 +02:00
William Lallemand	ccc7ee45f9	MINOR: httpclient: enable request buffering The request buffering is required for doing l7 retry. The IO handler of the httpclient need to be rework for that. This patch change the IO handler so it copies partially the data instead of swapping buffer. This is needed because the b_xfer won't never work if the destination buffer is not empty, which is the case when buffering.	2022-04-06 11:43:01 +02:00
Remi Tricot-Le Breton	e8041fe8bc	BUG/MINOR: ssl/cli: Remove empty lines from CLI output There were empty lines in the output of "show ssl ca-file <cafile>" and "show ssl crl-file <crlfile>" commands when an empty line should only mark the end of the output. This patch adds a space to those lines. This patch should be backported to 2.5.	2022-04-05 16:53:37 +02:00
William Lallemand	80296b4bd5	BUG/MINOR: ssl: handle X509_get_default_cert_dir() returning NULL ssl_store_load_locations_file() is using X509_get_default_cert_dir() when using '@system-ca' as a parameter. This function could return a NULL if OpenSSL was built with a X509_CERT_DIR set to NULL, this is uncommon but let's fix this. No backport needed, 2.6 only. Fix issue #1637.	2022-04-05 10:19:30 +02:00
Nikola Sale	0dbf03871f	MINOR: sample: converter: Add add_item convertor This new converter is similar to the concat converter and can be used to build new variables made of a succession of other variables but the main difference is that it does the checks if adding a delimiter makes sense as wouldn't be the case if e.g the current input sample is empty. That situation would require 2 separate rules using concat converter where the first rule would have to check if the current sample string is empty before adding a delimiter. This resolves GitHub Issue #1621.	2022-04-04 07:30:58 +02:00
William Lallemand	c6b1763dcd	MINOR: ssl: ca-file @system-ca loads the system trusted CA The new parameter "@system-ca" to the ca-file directives loads the trusted CA in the directory returned by X509_get_default_cert_dir().	2022-04-01 23:52:50 +02:00
William Lallemand	4f6ca32217	BUG/MINOR: ssl: continue upon error when opening a directory w/ ca-file Previous patch was accidentaly breaking upon an error when itarating through a CA directory. This is not the expected behavior, the function must start processing the other files after the warning.	2022-04-01 23:52:50 +02:00
William Lallemand	87fd994727	MEDIUM: ssl: allow loading of a directory with the ca-file directive This patch implements the ability to load a certificate directory with the "ca-file" directive. The X509_STORE_load_locations() API does not allow to cache a directory in memory at startup, it only references the directory to allow a lookup of the files when needed. But that is not compatible with the way HAProxy works, without any access to the filesystem. The current implementation loads every ".pem", ".crt", ".cer", and ".crl" available in the directory which is what is done when using c_rehash and X509_STORE_load_locations(). Those files are cached in the same X509_STORE referenced by the directory name. When looking at "show ssl ca-file", everything will be shown in the same entry. This will eventually allow to load more easily the CA of the system, which could already be done with "ca-file /etc/ssl/certs" in the configuration. Loading failure intentionally emit a warning instead of an alert, letting HAProxy starts when one of the files can't be loaded. Known limitations: - There is a bug in "show ssl ca-file", once the buffer is full, the iohandler is not called again to output the next entries. - The CLI API is kind of limited with this, since it does not allow to add or remove a entry in a particular ca-file. And with a lot of CAs you can't push them all in a buffer. It probably needs a "add ssl ca-file" like its done with the crt-list. Fix issue #1476.	2022-04-01 20:36:38 +02:00
Willy Tarreau	fcce0e1e09	OPTIM: hpack: read 32 bits at once when possible. As suggested in the comment, it's possible to read 32 bits at once in big-endian order, and now we have the functions to do this, so let's do it. This reduces the code on the fast path by 31 bytes on x86, and more importantly performs single-operation 32-bit reads instead of playing with shifts and additions.	2022-04-01 17:29:06 +02:00
Willy Tarreau	17b4687a89	CLEANUP: hpack: be careful about integer promotion from uint8_t As reported in issue #1635, there's a subtle sign change when shifting a uint8_t value to the left because integer promotion first turns any smaller type to signed int even if it was unsigned. A warning was reported about uint8_t shifted left 24 bits that couldn't fit in int due to this. It was verified that the emitted code didn't change, as expected, but at least this allows to silence the code checkers. There's no need to backport this.	2022-04-01 17:29:06 +02:00
Frédéric Lécaille	eb2a2da67c	BUG/MINOR: quic: Missing TX packet deallocations Ensure all TX packets are deallocated. There may be remaining ones which will never be acknowledged or deemed lost.	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	64670884ba	BUG/MINOR: quic: Missing ACK range deallocations free_quic_arngs() was implemented but not used. Let's call it from quic_conn_release().	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	96fd1633e1	BUG/MINOR: quic: QUIC TLS secrets memory leak We deallocate these secrets from quic_conn_release().	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	000162e1d6	BUG/MINOR: h3: Missing wait event struct field initialization This one has been detected by valgrind: ==2179331== Conditional jump or move depends on uninitialised value(s) ==2179331== at 0x1B6EDE: qcs_notify_recv (mux_quic.c:201) ==2179331== by 0x1A17C5: qc_handle_uni_strm_frm (xprt_quic.c:2254) ==2179331== by 0x1A1982: qc_handle_strm_frm (xprt_quic.c:2286) ==2179331== by 0x1A2CDB: qc_parse_pkt_frms (xprt_quic.c:2550) ==2179331== by 0x1A6068: qc_treat_rx_pkts (xprt_quic.c:3463) ==2179331== by 0x1A6C3D: quic_conn_app_io_cb (xprt_quic.c:3589) ==2179331== by 0x3AA566: run_tasks_from_lists (task.c:580) ==2179331== by 0x3AB197: process_runnable_tasks (task.c:883) ==2179331== by 0x357E56: run_poll_loop (haproxy.c:2750) ==2179331== by 0x358366: run_thread_poll_loop (haproxy.c:2921) ==2179331== by 0x3598D2: main (haproxy.c:3538) ==2179331==	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	b823bb7f7f	MINOR: quic: Add traces about list of frames This should be useful to have an idea of the list of frames which could be built towards the list of available frames when building packets. Same thing about the frames which could not be built because of a lack of room in the TX buffer.	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	6c01b74ffa	MINOR: quic: Useless call to SSL_CTX_set_default_verify_paths() This call to SSL_CTX_set_default_verify_paths() is useless for haproxy.	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	12fd259363	BUG/MINOR: quic: Too much prepared retransmissions due to anti-amplification We must not re-enqueue frames if we can detect in advance they will not be transmitted due to the anti-amplification limit.	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	009016c0cd	BUG/MINOR: quic: Non duplicated frames upon fast retransmission We must duplicate the frames to be sent again from packets which are not deemed lost.	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	5cfb4edca7	BUG/MINOR: quic: Do not probe from an already probing packet number space During a handshake, after having prepared a probe upon a PTO expiration from process_timer(), we wake up the I/O handler to make it send probing packets. This handler first treat incoming packets which trigger a fast retransmission leading to send too much probing (duplicated) packets. In this cas we cancel the fast retranmission.	2022-04-01 16:26:05 +02:00
Frédéric Lécaille	03235d78ae	MINOR: quic: Do not display any timer value from process_timer() This is confusing to display the connection timer from this function as it is not supposed to update it. Only qc_set_timer() should do that.	2022-04-01 16:22:52 +02:00
Frédéric Lécaille	05bd92bbc5	BUG/MINOR: quic: Discard Initial packet number space only one time When discarding a packet number space, we at least reset the PTO backoff counter. Doing this several times have an impact on the PTO duration calculation. We must not discard a packet number space several times (this is already the case for the handshake packet number space).	2022-04-01 16:22:52 +02:00
Frédéric Lécaille	d6570e1789	BUG/MINOR: quic: Missing probing packets when coalescing Before having a look at the next encryption level to build packets if there is no more ack-eliciting frames to send we must check we have not to probe from the current encryption level anymore. If not, we only send one datagram instead of sending two datagrams giving less chance to recover from packet loss.	2022-04-01 16:22:52 +02:00
Frédéric Lécaille	b002145e9f	MEDIUM: quic: Send ACK frames asap Due to a erroneous interpretation of the RFC 9000 (quic-transport), ACKs frames were always sent only after having received two ack-eliciting packets. This could trigger useless retransmissions for tail packets on the peer side. For now on, we send as soon as possible ACK frames as soon as we have ACK to send, in the same packets as the ack-eliciting frame packets, and we also send ACK frames after having received 2 ack-eliciting packets since the last time we sent an ACK frame with other ack-eliciting frames.	2022-04-01 16:22:52 +02:00
Frédéric Lécaille	205e4f359e	CLEANUP: quic: Remove all atomic operations on packet number spaces As such variables are handled by the QUIC connection I/O handler which runs always on the thread, there is no need to continue to use such atomic operations	2022-04-01 16:22:47 +02:00
Frédéric Lécaille	fc79006c92	CLEANUP: quic: Remove all atomic operations on quic_conn struct As the QUIC connections are always handled by the same thread there is no need anymore to continue to use atomic operations on such variables.	2022-04-01 16:22:44 +02:00
Frédéric Lécaille	f44d19eb91	BUG/MEDIUM: quic: Possible crash in ha_quic_set_encryption_secrets() This bug has come with this commit: `1fc5e16c4` MINOR: quic: More accurate immediately close As mentionned in this commit we do not want to derive anymore secret when in closing state. But the flag which denote secrets were derived was set. Add a label at the correct flag to skip the secrets derivation without setting this flag.	2022-04-01 16:22:40 +02:00
Willy Tarreau	413713f02a	BUG/MAJOR: mux_pt: always report the connection error to the conn_stream Over time we've tried hard to abstract connection errors from the upper layers so that they're reported per stream and not per connection. As early as 1.8-rc1, commit `4ff3b8964` ("MINOR: connection: make conn_stream users also check for per-stream error flag") did precisely this, but strangely only for rx, not for tx (probably that by then send errors were not imagined to be reported that way). And this lack of Tx error check was just revealed in 2.6 by recent commit `d1480cc8a` ("BUG/MEDIUM: stream-int: do not rely on the connection error once established") that causes wakeup loops between si_cs_send() failing to send via mux_pt_snd_buf() and subscribing against si_cs_io_cb() in loops because the function now rightfully only checks for CS_FL_ERROR and not CO_FL_ERROR. As found by Amaury, this causes aborted "show events -w" to cause haproxy to loop at 100% CPU. This fix theoretically needs to be backported to all versions, though it will be necessary and sufficient to backport it wherever `4ff3b8964` gets backported.	2022-03-31 16:55:52 +02:00
Willy Tarreau	c40c407268	BUG/MINOR: cli/stream: fix "shutdown session" to iterate over all threads The list of streams was modified in 2.4 to become per-thread with commit `a698eb673` ("MINOR: streams: use one list per stream instead of a global one"). However the change applied to cli_parse_shutdown_session() is wrong, as it uses the nullity of the stream pointer to continue on next threads, but this one is not null once the list_for_each_entry() loop is finished, it points to the list's head again, so the loop doesn't check other threads, and no message is printed either to say that the stream was not found. Instead we should check if the stream is equal to the requested pointer since this is the condition to break out of the loop. Thus must be backported to 2.4. Thanks to Maciej Zdeb for reporting this.	2022-03-31 15:00:45 +02:00
Amaury Denoyelle	d8e680cbaf	MEDIUM: mux-quic: remove qcs tree node The new qc_stream_desc type has a tree node for storage. Thus, we can remove the node in the qcs structure. When initializing a new stream, it is stored into the qcc streams_by_id tree. When the MUX releases it, it will freed as soon as its buffer is emptied. Before this, the quic-conn is responsible to store it inside its own streams_by_id tree.	2022-03-30 16:26:59 +02:00
Amaury Denoyelle	7272cd76fc	MEDIUM: quic: move transport fields from qcs to qc_conn_stream Move the xprt-buf and ack related fields from qcs to the qc_stream_desc structure. In exchange, qcs has a pointer to the low-level stream. For each new qcs, a qc_stream_desc is automatically allocated. This simplify the transport layer by removing qcs/mux manipulation during ACK frame parsing. An additional check is done to not notify the MUX on sending if the stream is already released : this case may now happen on retransmission. To complete this change, the quic_stream frame now references the quic_stream instance instead of a qcs.	2022-03-30 16:19:48 +02:00
Amaury Denoyelle	5c3859c509	MINOR: quic: implement stream descriptor for transport layer Currently, the mux qcs streams manage the Tx buffering, even after sending it to the transport layer. Buffers are emptied when acknowledgement are treated by the transport layer. This complicates the MUX liberation and we may loose some data after the MUX free. Change this paradigm by moving the buffering on the transport layer. For this goal, a new type is implemented as low-level stream at the transport layer, as a counterpart of qcs mux instances. This structure is called qc_stream_desc. This will allow to free the qcs/qcc instances without having to wait for acknowledge reception. For the moment, the quic-conn is responsible to store the qc_stream_desc in a new tree named streams_by_id. This will sligthly change in the next commits to remove the qcs node which has a similar purpose : qc_stream_desc instances will be shared between the qcc MUX and the quic-conn. This patch only introduces the new type definition and the function to manipulate it. The following commit will bring the rearchitecture in the qcs structure.	2022-03-30 16:16:07 +02:00
Amaury Denoyelle	95e50fbeff	CLEANUP: quic: complete comment on qcs_try_to_consume Specify the return value usage.	2022-03-30 16:12:18 +02:00
Amaury Denoyelle	f89094510c	BUG/MINOR: mux-quic: ensure to free all qcs on MUX release Remove qcs instances left during qcc MUX release. This can happen when the MUX is closed before the completion of all the transfers, such as on a timeout or process termination. This may free some memory leaks on the connection.	2022-03-30 16:12:18 +02:00
Amaury Denoyelle	8347f27221	BUG/MINOR: h3: release resources on close Implement the release app-ops ops for H3 layer. This is used to clean up uni-directional streams and the h3 context. This prevents a memory leak on H3 resources for each connection.	2022-03-30 16:12:18 +02:00
Amaury Denoyelle	cbc13b71c6	MINOR: mux-quic: define release app-ops Define a new callback release inside qcc_app_ops. It is called when the qcc MUX is freed via qc_release. This will allows to implement cleaning on the app layer.	2022-03-30 16:12:18 +02:00
Amaury Denoyelle	dccbd733f0	MINOR: mux-quic: reorganize qcs free Regroup some cleaning operations inside a new function qcs_free. This can be used for all streams, both through qcs_destroy and with uni-directional streams.	2022-03-30 16:12:18 +02:00
Amaury Denoyelle	50742294f5	MINOR: mux-quic: return qcs instance from qcc_get_qcs Refactoring on qcc_get_qcs : return the qcs instance instead of the tree node. This is useful to hide some eb64_entry macros for better readability.	2022-03-30 16:12:18 +02:00
Amaury Denoyelle	8d5def0bab	BUG/MEDIUM: quic: do not use qcs from quic_stream on ACK parsing The quic_stream frame stores the qcs instance. On ACK parsing, qcs is accessed to clear the stream buffer. This can cause a segfault if the MUX or the qcs is already released. Consider the following scenario : 1. a STREAM frame is generated by the MUX transport layer emits the frame with PKN=1 upper layer has finished the transfer so related qcs is detached 2. transport layer reemits the frame with PKN=2 because ACK was not received 3. ACK for PKN=1 is received, stream buffer is cleared at this stage, qcs may be freed by the MUX as it is detached 4. ACK for PKN=2 is received qcs for STREAM frame is dereferenced which will lead to a crash To prevent this, qcs is never accessed from the quic_stream during ACK parsing. Instead, a lookup is done on the MUX streams tree. If the MUX is already released, no lookup is done. These checks prevents a possible segfault. This change may have an impact on the perf as now we are forced to use a tree lookup operation. If this is the case, an alternative solution may be to implement a refcount on qcs instances.	2022-03-30 16:12:18 +02:00
William Lallemand	5213918fe2	BUILD: ssl/lua: CacheCert needs OpenSSL Return an lua error when trying to use CacheCert.set() and OpenSSL was not used to build HAProxy.	2022-03-30 15:05:42 +02:00
William Lallemand	30fcca18a5	MINOR: ssl/lua: CertCache.set() allows to update an SSL certificate file The CertCache.set() function allows to update an SSL certificate file stored in the memory of the HAProxy process. This function does the same as "set ssl cert" + "commit ssl cert" over the CLI. This could be used to update the crt and key, as well as the OCSP, the SCTL, and the OSCP issuer. The implementation does yield every 10 ckch instances, the same way the "commit ssl cert" do.	2022-03-30 14:56:10 +02:00
William Lallemand	26654e7a59	MINOR: ssl: add "crt" in the cert_exts array The cert_exts array does handle "crt" the default way, however you might stil want to look for these extensions in the array.	2022-03-30 14:55:53 +02:00
William Lallemand	e60c7d6e59	MINOR: ssl: export ckch_inst_rebuild() ckch_inst_rebuild() will be needed to regenerate the ckch instances from the lua code, we need to export it.	2022-03-30 12:18:16 +02:00
William Lallemand	ff8bf988b9	MINOR: ssl: simplify the certificate extensions array Simplify the "cert_exts" array which is used for the selection of the parsing function depending on the extension. It now uses a pointer to an array element instead of an index, which is simplier for the declaration of the array. This way also allows to have multiple extension using the same type.	2022-03-30 12:18:16 +02:00
William Lallemand	aaacc7e8ad	MINOR: ssl: move the cert_exts and the CERT_TYPE enum Move the cert_exts declaration and the CERT_TYPE enum in the .h in order to reuse them in another file.	2022-03-30 12:18:16 +02:00
William Lallemand	3b5a3a6c03	MINOR: ssl: split the cert commit io handler Extract the code that replace the ckch_store and its dependencies into the ckch_store_replace() function. This function must be used under the global ckch lock. It frees everything related to the old ckch_store.	2022-03-30 12:18:16 +02:00
William Lallemand	7177a95f89	MEDIUM: httpclient/lua: be stricter with httpclient parameters Checks the argument passed to the httpclient send functions so we don't add a mispelled argument.	2022-03-30 12:18:16 +02:00
Willy Tarreau	3f0b2e85f3	MINOR: services: alphabetically sort service names Note that we cannot reuse dump_act_rules() because the output format may be adjusted depending on the call place (this is also used from haproxy -vv). The principle is the same however.	2022-03-30 12:12:44 +02:00
Willy Tarreau	0f99637353	MINOR: filters: alphabetically sort the list of filter names There are very few but they're registered from constructors, hence in a random order. The scope had to be copied when retrieving the next keyword. Note that this also has the effect of listing them sorted in haproxy -vv.	2022-03-30 12:08:00 +02:00
Willy Tarreau	4465171593	MINOR: cli: alphanumerically sort the dump of supported commands Like for previous keyword classes, we're sorting the output. But this time as it's not trivial to do it with multiple words, instead we're proceeding like the help command, we sort them on their usage message when present, and fall back to the first word of the command when there is no usage message (e.g. "help" command).	2022-03-30 12:02:35 +02:00
Willy Tarreau	800bd78944	MINOR: acl: alphanumerically sort the ACL dump The mechanism is similar to others. We take care of sorting on the keyword only and not the fetch_kw which is not unique.	2022-03-30 11:49:59 +02:00
Willy Tarreau	5e0f90a930	MINOR: sample: alphanumerically sort sample & conv keyword dumps It's much more convenient to sort these keywords on output to detect changes, and it's easy to do. The patch looks big but most of it is only caused by an indent change in the loop, as "git diff -b" shows.	2022-03-30 11:30:36 +02:00
Willy Tarreau	d905826000	MINOR: config: alphanumerically sort config keywords output The output produced by dump_registered_keywords() really deserves to be sorted in order to ease comparisons. The function now implements a tiny sorting mechanism that's suitable for each two-level list, and makes use of dump_act_rules() to dump rulesets. The code is not significantly more complicated and some parts (e.g options) could even be factored. The output is much more exploitable to detect differences now.	2022-03-30 11:21:32 +02:00
Willy Tarreau	2100b383ab	MINOR: action: add a function to dump the list of actions for a ruleset The new function dump_act_rules() now dumps the list of actions supported by a ruleset. These actions are alphanumerically sorted first so that the produced output is easy to compare.	2022-03-30 11:19:22 +02:00
Willy Tarreau	3ff476e9ef	MINOR: tools: add strordered() to check whether strings are ordered When trying to sort sets of strings, it's often needed to required to compare 3 strings to see if the chosen one fits well between the two others. That's what this function does, in addition to being able to ignore extremities when they're NULL (typically for the first iteration for example).	2022-03-30 10:02:56 +02:00
Willy Tarreau	29d799d591	MINOR: sample: list registered sample converter functions Similar to the sample fetch keywords, let's also list the converter keywords. They're much simpler since there's no compatibility matrix. Instead the input and output types are listed. This is called by dump_registered_keywords() for the "cnv" keywords class.	2022-03-29 18:01:37 +02:00
Willy Tarreau	f78813f74f	MINOR: samples: add a function to list register sample fetch keywords New function smp_dump_fetch_kw lists registered sample fetch keywords with their compatibility matrix, mandatory and optional argument types, and output types. It's called from dump_registered_keywords() with class "smp".	2022-03-29 18:01:37 +02:00
Willy Tarreau	6ff7d1b9a5	MINOR: acl: add a function to dump the list of known ACL keywords New function acl_dump_kwd() dumps the registered ACL keywords and their sample-fetch equivalent to stdout. It's called by dump_registered_keywords() for keyword class "acl".	2022-03-29 18:01:37 +02:00
Willy Tarreau	06d0e2e034	MINOR: cli: add a new keyword dump function New function cli_list_keywords() scans the list of registered CLI keywords and dumps them on stdout. It's now called from dump_registered_keywords() for the class "cli". Some keywords are valid for the master, they'll be suffixed with "[MASTER]". Others are valid for the worker, they'll have "[WORKER]". Those accessible only in expert mode will show "[EXPERT]" and the experimental ones will show "[EXPERIM]".	2022-03-29 18:01:37 +02:00
Willy Tarreau	5fcc100d91	MINOR: services: extend list_services() to dump to stdout When no output stream is passed, stdout is used with one entry per line, and this is called from dump_registered_services() when passed the class "svc".	2022-03-29 18:01:37 +02:00
Willy Tarreau	3b65e14842	MINOR: filters: extend flt_dump_kws() to dump to stdout When passing a NULL output buffer the function will now dump to stdout with a more compact format that is more suitable for machine processing. An entry was added to dump_registered_keyword() to call it when the keyword class "flt" is requested.	2022-03-29 18:01:37 +02:00
Willy Tarreau	ca1acd6080	MINOR: config: add a function to dump all known config keywords All registered config keywords that are valid in the config parser are dumped to stdout organized like the regular sections (global, listen, etc). Some keywords that are known to only be valid in frontends or backends will be suffixed with [FE] or [BE]. All regularly registered "bind" and "server" keywords are also dumped, one per "bind" or "server" line. Those depending on ssl are listed after the "ssl" keyword. Doing so required to export the listener and server keyword lists that were static. The function is called from dump_registered_keywords() for keyword class "cfg".	2022-03-29 18:01:32 +02:00
Willy Tarreau	76871a4f8c	MINOR: management: add some basic keyword dump infrastructure It's difficult from outside haproxy to detect the supported keywords and syntax. Interestingly, many of our modern keywords are enumerated since they're registered from constructors, so it's not very hard to enumerate most of them. This patch creates some basic infrastructure to support dumping existing keywords from different classes on stdout. The format will differ depending on the classes, but the idea is that the output could easily be passed to a script that generates some simple syntax highlighting rules, completion rules for editors, syntax checkers or config parsers. The principle chosen here is that if "-dK" is passed on the command-line, at the end of the parsing the registered keywords will be dumped for the requested classes passed after "-dK". Special name "help" will show known classes, while "all" will execute all of them. The reason for doing that after the end of the config processor is that it will also enumerate internally-generated keywords, Lua or even those loaded from external code (e.g. if an add-on is loaded using LD_PRELOAD). A typical way to call this with a valid config would be: ./haproxy -dKall -q -c -f /path/to/config If there's no config available, feeding /dev/null will also do the job, though it will not be able to detect dynamically created keywords, of course. This patch also updates the management doc. For now nothing but the help is listed, various subsystems will follow in subsequent patches.	2022-03-29 17:55:54 +02:00
Willy Tarreau	0d71d2f4fa	BUG/MINOR: samples: add missing context names for sample fetch functions In 2.4, two commits added support for supporting sample fetch calls from new config and CLI contexts, but these were not added to the visibile names, which may possibly cause "(null)" to appear in some error messages. The commit in question were: `db5e0dbea` ("MINOR: sample: add a new CLI_PARSER context for samples") `f9a7a8fd8` ("MINOR: sample: add a new CFG_PARSER context for samples") This patch needs to be backported where these are present (2.4 and above).	2022-03-29 17:55:54 +02:00
Christopher Faulet	b4f96eda56	BUG/MINOR: log: Initialize the list element when allocating a new log server `211ea252d` ("BUG/MINOR: logs: fix logsrv leaks on clean exit") introduced a regression because the list element of a new log server is not intialized. Thus HAProxy crashes on error path when an invalid log server is released. This patch shoud fix the issue #1636. It must be backported if the above commit is backported. For now, it is 2.6-specific and no backport is needed.	2022-03-29 14:17:10 +02:00
Christopher Faulet	744451c7c4	BUG/MEDIUM: mux-h1: Properly detect full buffer cases during message parsing When the destination buffer is full while there are still data to parse, the h1s must be marked as congested to be able to restart the parsing later. This work on headers and data parsing. But on trailers parsing, we fail to do so when the buffer is full before to parse the trailers. In this case, we skip the trailers parsing but the h1s is not marked as congested. This is important to be sure to wake up the mux to restart the parsing when some room is made in the buffer. Because of this bug, the message processing may hang till a timeout is triggered. Note that for 2.3 and 2.2, the EOM processing is buggy too, for the same reason. It should be fixed too on these versions. On the 2.0, only trailers parsing is affected. This patch must be backported as far as 2.0. On 2.3 and 2.2, the EOM parsing must be fixed too.	2022-03-28 16:19:04 +02:00
Christopher Faulet	d9fc12818e	BUG/MEDIUM: mux-fcgi: Properly handle return value of headers/trailers parsing h1_parse_msg_hdrs() and h1_parse_msg_tlrs() may return negative values if the parsing fails or if more space is needed in the destination buffer. When h1-htx was changed, The H1 mux was updated accordingly but not the FCGI mux. Thus if a negative value is returned, it is ignored and it is casted to a size_t, leading to an integer overflow on the <ofs> value, used to know the position in the RX buffer. This patch must be backported as far as 2.2.	2022-03-28 15:56:17 +02:00
William Lallemand	3d7a9186dd	BUG/MINOR: tools: url2sa reads too far when no port nor path url2sa() still have an unfortunate case where it reads 1 byte too far, it happens when no port or path are specified in the URL, and could crash if the byte after the URL is not allocated (mostly with ASAN). This case is never triggered in old versions of haproxy because url2sa is used with buffers which are way bigger than the URL. It is only triggered with the httpclient. Should be bacported in every stable branches.	2022-03-25 17:48:28 +01:00
Amaury Denoyelle	d8769d1d87	CLEANUP: h3: suppress by default stdout traces H3_DEBUG definition is removed from h3.c similarly to the commit `d96361b270` CLEANUP: qpack: suppress by default stdout traces Also, a plain fprintf in h3_snd_buf has been replaced to be conditional to the H3_DEBUG definition. These changes reduces the default output on stdout with QUIC traffic.	2022-03-25 15:30:23 +01:00
Amaury Denoyelle	d96361b270	CLEANUP: qpack: suppress by default stdout traces Remove the definition of DEBUG_HPACK on qpack-dec.c which forces the QPACK decoding traces on stderr. Also change the name to use a dedicated one for QPACK decoding as DEBUG_QPACK.	2022-03-25 15:22:40 +01:00
Amaury Denoyelle	18a10d07b6	BUILD: qpack: fix unused value when not using DEBUG_HPACK If the macro is not defined, some local variables are flagged as unused by the compiler. Fix this by using the __maybe_unused attribute. For now, the macro is defined in the qpack-dec.c. However, this will change to not mess up the stderr output of haproxy with QUIC traffic.	2022-03-25 15:21:45 +01:00
Amaury Denoyelle	251eadfce5	MINOR: mux-quic: activate qmux traces on stdout via macro This commit is similar to the following one : commit `118b2cbf84` MINOR: quic: activate QUIC traces at compilation If the macro ENABLE_QUIC_STDOUT_TRACES is defined, qmux traces are outputted automatically on stdout. This is useful for the haproxy-qns interop docker image.	2022-03-25 14:51:14 +01:00
Amaury Denoyelle	fdcec3644a	MINOR: mux-quic: add trace event for qcs_push_frame Add a new qmux trace event QMUX_EV_QCS_PUSH_FRM. Its only purpose is to display the meaningful result of a qcs_push_frame invocation. A dedicated struct qcs_push_frm_trace_arg is defined to pass a series of extra args for the trace output.	2022-03-25 14:51:14 +01:00
Amaury Denoyelle	fa29f33f2c	MINOR: mux-quic: add trace event for frame sending Define a new qmux event QMUX_EV_SEND_FRM. This allows to pass a quic_frame as an extra argument. Depending on the frame type, a special format can be used to log the frame content. Currently this event is only used in qc_send_max_streams. Thus the handler is only able to handle MAX_STREAMS frames.	2022-03-25 14:51:14 +01:00
Amaury Denoyelle	4f137577d7	MINOR: mux-quic: replace printfs by traces Convert all printfs in the mux-quic code with traces. Note that some meaningul printfs were not converted because they use extra args in a format-string. This is the case inside qcs_push_frame and qc_send_max_streams. A dedicated trace event should be implemented for them to be able to display the extra arguments.	2022-03-25 14:51:14 +01:00
Amaury Denoyelle	dd4fbfb5b4	MINOR: mux-quic: declare the qmux trace module Declare a new trace module for mux-quic named qmux. It will be used to convert all printf to regular traces. The handler qmux_trace can uses a connection and a qcs instance as extra arguments.	2022-03-25 14:51:10 +01:00
Amaury Denoyelle	0c2d964280	REORG: quic: use a dedicated quic_loss.c Move all inline functions with trace from quic_loss.h to a dedicated object file. This let to remove the TRACE_SOURCE macro definition outside of the include file. This change is required to be able to define another TRACE_SOUCE inside the mux_quic.c for a dedicated trace module.	2022-03-25 14:45:45 +01:00
Amaury Denoyelle	777969c163	BUILD: quic: add missing includes Complete the include list for the files quic_loss.h and quic_sock.c.	2022-03-25 14:45:45 +01:00
Amaury Denoyelle	9296091cf7	MINOR: mux-quic: convert fin on push-frame as boolean This is only useful to display a clear 0/1 value in the traces. This has no impact beyond this cosmetic change.	2022-03-25 14:45:45 +01:00
William Lallemand	b938b77ade	BUG/MINOR: tools: fix url2sa return value with IPv4 Fix `8a91374` ("BUG/MINOR: tools: url2sa reads ipv4 too far") introduced a regression in the value returned when parsing an ipv4 host. Tthe consumed length is supposed to be as far as the first character of the path, only its not computed correctly anymore and return the length minus the size of the scheme. Fixed the issue by reverting 'curr' and 'url' as they were before the patch. Must be backported in every stable branch where the `8a91374` patch was backported.	2022-03-25 11:49:27 +01:00
Frédéric Lécaille	cc2764e7fe	BUG/MINOR: quic: Wrong buffer length passed to generate_retry_token() After having consumed <i> bytes from <buf>, the remaining available room to be passed to generate_retry_token() is sizeof(buf) - i. This bug could be easily reproduced with quic-qo as client which chooses a random value as ODCID length.	2022-03-23 17:16:20 +01:00
Willy Tarreau	0c3205a541	BUILD: stream-int: avoid a build warning when DEBUG is empty When no DEBUG_STRICT is enabled, we get this build warning: src/stream_interface.c: In function 'stream_int_chk_snd_conn': src/stream_interface.c:1198:28: warning: unused variable 'conn' [-Wunused-variable] 1198 \| struct connection *conn = cs_conn(cs); \| ^~~~ This was the result of the simplification of the code in commit `d1480cc8a` ("BUG/MEDIUM: stream-int: do not rely on the connection error once established") which removed the last user of this variable outside of a BUG_ON(). If the patch above is backported, this one should be backported as well.	2022-03-23 11:15:49 +01:00
Amaury Denoyelle	1e5e5136ee	MINOR: mux-quic: support MAX_DATA frame parsing This commit is similar to the previous one but with MAX_DATA frames. This allows to increase the connection level flow-control limit. If the connection was blocked due to QC_CF_BLK_MFCTL flag, the flag is reseted.	2022-03-23 10:14:14 +01:00
Amaury Denoyelle	8727ff4668	MINOR: mux-quic: support MAX_STREAM_DATA frame parsing Implement a MUX method to parse MAX_STREAM_DATA. If the limit is greater than the previous one and the stream was blocked, the flag QC_SF_BLK_SFCTL is removed.	2022-03-23 10:09:39 +01:00
Amaury Denoyelle	05ce55e582	MEDIUM: mux-quic: respect peer connection data limit This commit is similar to the previous one, but this time on the connection level instead of the stream. When the connection limit is reached, the connection is flagged with QC_CF_BLK_MFCTL. This flag is checked in qc_send. qcs_push_frame uses a new parameter which is used to not exceed the connection flow-limit while calling it repeatdly over multiple streams instance before transfering data to the transport layer.	2022-03-23 10:05:29 +01:00
Amaury Denoyelle	6ea781919a	MEDIUM: mux-quic: respect peer bidirectional stream data limit Implement the flow-control max-streams-data limit on emission. We ensure that we never push more than the offset limit set by the peer. When the limit is reached, the stream is marked as blocked with a new flag QC_SF_BLK_SFCTL to disable emission. Currently, this is only implemented for bidirectional streams. It's required to unify the sending for unidirectional streams via qcs_push_frame from the H3 layer to respect the flow-control limit for them.	2022-03-23 10:05:29 +01:00
Amaury Denoyelle	78396e5ee8	MINOR: mux-quic: use shorter name for flow-control fields Rename the fields used for flow-control in the qcc structure. The objective is to have shorter name for better readability while keeping their purpose clear. It will be useful when the flow-control will be extended with new fields.	2022-03-23 10:05:29 +01:00
Amaury Denoyelle	75d14ad5cb	MINOR: mux-quic: add comments for send functions Add comments on qc_send and qcs_push_frame. Also adjust the return of qc_send to reflect the total bytes sent. This has no impact as currently the return value is not checked by the caller.	2022-03-23 10:05:29 +01:00
Amaury Denoyelle	ac74aa531d	MINOR: mux-quic: complete trace when stream is not found Display the ID of the stream not found. This will help to detect when we received retransmitted frames for an already closed stream.	2022-03-23 09:49:08 +01:00
Amaury Denoyelle	e0320b8aa6	CLEANUP: mux-quic: change comment style to not mess with git conflict Remove "=======" symbols from the MUX buffer diagram. This is useful to not mess with git conflict markers when resolving a conflict.	2022-03-23 09:48:43 +01:00
Frédéric Lécaille	aaf1f19e8b	MINOR: quic: Add traces in qc_set_timer() (scheduling) This should be helpful to diagnose some issues: timer task not run when it should run.	2022-03-23 09:01:45 +01:00
Frédéric Lécaille	ce69cbc520	MINOR: quic: Add traces about stream TX buffer consumption This will be helpful to diagnose STREAM blocking states.	2022-03-23 09:01:45 +01:00
Dhruv Jain	1295798139	MEDIUM: mqtt: support mqtt_is_valid and mqtt_field_value converters for MQTTv3.1 In MQTTv3.1, protocol name is "MQIsdp" and protocol level is 3. The mqtt converters(mqtt_is_valid and mqtt_field_value) did not work for clients on mqttv3.1 because the mqtt_parse_connect() marked the CONNECT message invalid if either the protocol name is not "MQTT" or the protocol version is other than v3.1.1 or v5.0. To fix it, we have added the mqttv3.1 protocol name and version as part of the checks. This patch fixes the mqtt converters to support mqttv3.1 clients as well (issue #1600). It must be backported to 2.4.	2022-03-22 09:25:52 +01:00
Frédéric Lécaille	411aa6daf5	BUG/MINOR: quic: Non initialized variable in quic_build_post_handshake_frames() <cid> could be accessed before being initialized.	2022-03-21 14:30:23 +01:00
Frédéric Lécaille	44ae75220a	BUG/MINOR: quic: Incorrect peer address validation We must consider the peer address as validated as soon as we received an handshake packet. An ACK frame in handshake packet was too restrictive. Rename the concerned flag to reflect this situation.	2022-03-21 14:27:09 +01:00
Frédéric Lécaille	12aa26b6fd	BUG/MINOR: quic: 1RTT packets ignored after mux was released We must be able to handle 1RTT packets after the mux has terminated its job (qc->mux_state == QC_MUX_RELEASED). So the condition (qc->mux_state != QC_MUX_READY) in qc_qel_may_rm_hp() is not correct when we want to wait for the mux to be started. Add a check in qc_parse_pkt_frms() to ensure is started before calling it. All the STREAM frames will be ignored when the mux will be released.	2022-03-21 14:27:09 +01:00
Frédéric Lécaille	2899fe2460	BUG/MINOR: quic: Missing TX packet initializations The most important one is the ->flags member which leads to an erratic xprt behavior. For instance a non ack-eliciting packet could be seen as ack-eliciting leading the xprt to try to retransmit a packet which are not ack-eliciting. In this case, the xprt does nothing and remains indefinitively in a blocking state.	2022-03-21 14:27:09 +01:00
Frédéric Lécaille	f27b66faee	BUG/MINOR: mux-quic: Missing I/O handler events initialization This could lead to a mux erratic behavior. Sometimes the application layer could not wakeup the mux I/O handler because it estimated it had already subscribed to write events (see h3_snd_buf() end of implementation).	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	4e22f28feb	BUG/MINOR: mux-quic: Access to empty frame list from qc_send_frames() This was revealed by libasan when each time qc_send_frames() is run at the first time: ================================================================= ==84177==ERROR: AddressSanitizer: stack-buffer-overflow on address 0x7fbaaca2b3c8 at pc 0x560a4fdb7c2e bp 0x7fbaaca2b300 sp 0x7fbaaca2b2f8 READ of size 1 at 0x7fbaaca2b3c8 thread T6 #0 0x560a4fdb7c2d in qc_send_frames src/mux_quic.c:473 #1 0x560a4fdb83be in qc_send src/mux_quic.c:563 #2 0x560a4fdb8a6e in qc_io_cb src/mux_quic.c:638 #3 0x560a502ab574 in run_tasks_from_lists src/task.c:580 #4 0x560a502ad589 in process_runnable_tasks src/task.c:883 #5 0x560a501e3c88 in run_poll_loop src/haproxy.c:2675 #6 0x560a501e4519 in run_thread_poll_loop src/haproxy.c:2846 #7 0x7fbabd120ea6 in start_thread nptl/pthread_create.c:477 #8 0x7fbabcb19dee in __clone (/lib/x86_64-linux-gnu/libc.so.6+0xfddee) Address 0x7fbaaca2b3c8 is located in stack of thread T6 at offset 56 in frame #0 0x560a4fdb7f00 in qc_send src/mux_quic.c:514 This frame has 1 object(s): [32, 48) 'frms' (line 515) <== Memory access at offset 56 overflows this variable HINT: this may be a false positive if your program uses some custom stack unwind mechanism, swapcontext or vfork (longjmp and C++ exceptions are supported) Thread T6 created by T0 here: #0 0x7fbabd1bd2a2 in __interceptor_pthread_create ../../../../src/libsanitizer/asan/asan_interceptors.cpp:214 #1 0x560a5036f9b8 in setup_extra_threads src/thread.c:221 #2 0x560a501e70fd in main src/haproxy.c:3457 #3 0x7fbabca42d09 in __libc_start_main ../csu/libc-start.c:308 SUMMARY: AddressSanitizer: stack-buffer-overflow src/mux_quic.c:473 in qc_send_frames	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	dcc74ff792	BUG/MINOR: quic: Unsent frame because of qc_build_frms() There are non already identified rare cases where qc_build_frms() does not manage to size frames to be encoded in a packet leading qc_build_frm() to fail to add such frame to the packet to be built. In such cases we must move back such frames to their origin frame list passed as parameter to qc_build_frms(): <frms>. because they were added to the packet frame list (but not built). If this this packet is not retransmitted, the frame is lost for ever! Furthermore we must not modify the buffer.	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	d64f68fb0a	BUG/MINOR: quic: Possible leak in quic_build_post_handshake_frames() Rework this function to leave the connection passed as parameter in the same state it was before entering this function.	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	f1f812bfdb	BUG/MINOR: quic: Possible crash in parse_retry_token() We must check the decoded length of this incoming data before copying into our internal structure. This could lead to crashes. Reproduced with such a packet captured from QUIC interop. { 0xc5, 0x00, 0x00, 0x00, 0x01, 0x12, 0xf2, 0x65, 0x4d, 0x9d, 0x58, 0x90, 0x23, 0x7e, 0x67, 0xef, 0xf8, 0xef, 0x5b, 0x87, 0x48, 0xbe, 0xde, 0x7a, /* corrupted byte: 0x11, */ 0x01, 0xdc, 0x41, 0xbf, 0xfb, 0x07, 0x39, 0x9f, 0xfd, 0x96, 0x67, 0x5f, 0x58, 0x03, 0x57, 0x74, 0xc7, 0x26, 0x00, 0x45, 0x25, 0xdc, 0x7f, 0xf1, 0x22, 0x1d, }	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	e2a1c1b372	MEDIUM: quic: Rework of the TX packets memory handling The TX packet refcounting had come with the multithreading support but not only. It is very useful to ease the management of the memory allocated for TX packets with TX frames attached to. At some locations of the code we have to move TX frames from a packet to a new one during retranmission when the packet has been deemed as lost or not. When deemed lost the memory allocated for the paquet must be released contrary to when its frames are retransmitted when probing (PTO). For now on, thanks to this patch we handle the TX packets memory this way. We increment the packet refcount when: - we insert it in its packet number space tree, - we attache an ack-eliciting frame to it. And reciprocally we decrement this refcount when: - we remove an ack-eliciting frame from the packet, - we delete the packet from its packet number space tree. Note that an optimization WOULD NOT be to fully reuse (without releasing its memorya TX packet to retransmit its contents (its ack-eliciting frames). Its information (timestamp, in flight length) to be processed by packet loss detection and the congestion control.	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	141982a4e1	MEDIUM: quic: Limit the number of ACK ranges When building a packet with an ACK frame, we store the largest acknowledged packet number sent in this frame in the packet (quic_tx_packet struc). When receiving an ack for such a packet we can purge the tree of acknowledged packet number ranges from the range sent before this largest acknowledged packet number.	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	8f3ae0272f	CLEANUP: quic: "largest_acked_pn" pktns struc member moving This struct member stores the largest acked packet number which was received. It is used to build (TX) packet. But this is confusing to store it in the tx packet of the packet number space structure even if it is used to build and transmit packets.	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	302c2b1120	MINOR: quic: Code factorization (TX buffer reuse) Add qc_may_reuse_cbuf() function used by qc_prep_pkts() and qc_prep_app_pkts(). Simplification of the factorized section code: there is no need to check there is enough room to mark the end of the data in the TX buf. This is done by the callers (qc_prep_pkts() and qc_prep_app_pkts()). Add a diagram to explain the conditions which must be verified to be able to reuse a cbuf struct. This should improve the QUIC stack implementation maintenability.	2022-03-21 11:29:40 +01:00
Tim Duesterhus	f4f6c0f6bb	CLEANUP: Reapply ist.cocci This makes use of the newly added: - i.ptr = p; - i.len = strlen(i.ptr); + i = ist(p); patch.	2022-03-21 08:30:47 +01:00
Tim Duesterhus	7750850594	CLEANUP: Reapply ist.cocci with `--include-headers-for-types --recursive-includes` Previous uses of `ist.cocci` did not add `--include-headers-for-types` and `--recursive-includes` preventing Coccinelle seeing `struct ist` members of other structs. Reapply the patch with proper flags to further clean up the use of the ist API. The command used was: spatch -sp_file dev/coccinelle/ist.cocci -in_place --include-headers --include-headers-for-types --recursive-includes --dir src/	2022-03-21 08:30:47 +01:00
Christopher Faulet	ab398d8ff9	BUG/MINOR: http-rules: Don't free new rule on allocation failure If allocation of a new HTTP rule fails, we must not release it calling free_act_rule(). The regression was introduced by the commit `dd7e6c6dc` ("BUG/MINOR: http-rules: completely free incorrect TCP rules on error"). This patch must only be backported if the commit above is backported. It should fix the issues #1627, #1628 and #1629.	2022-03-21 08:24:17 +01:00
Christopher Faulet	9075dbdd84	BUG/MINOR: rules: Initialize the list element when allocating a new rule `dd7e6c6dc` ("BUG/MINOR: http-rules: completely free incorrect TCP rules on error") and `388c0f2a6` ("BUG/MINOR: tcp-rules: completely free incorrect TCP rules on error") introduced a regression because the list element of a new rule is not intialized. Thus HAProxy crashes when an incorrect rule is released. This patch must be backported if above commits are backported. Note that new_act_rule() only exists since the 2.5. It relies on the commit `d535f807b` ("MINOR: rules: add a new function new_act_rule() to allocate act_rules").	2022-03-21 07:55:37 +01:00
Willy Tarreau	15a4733d5d	BUG/MEDIUM: mux-h2: make use of http-request and keep-alive timeouts Christian Ruppert reported an issue explaining that it's not possible to forcefully close H2 connections which do not receive requests anymore if they continue to send control traffic (window updates, ping etc). This will indeed refresh the timeout. In H1 we don't have this problem because any single byte is part of the stream, so the control frames in H2 would be equivalent to TCP acks in H1, that would not contribute to the timeout being refreshed. What misses from H2 is the use of http-request and keep-alive timeouts. These were not implemented because initially it was hard to see how they could map to H2. But if we consider the real use of the keep-alive timeout, that is, how long do we keep a connection alive with no request, then it's pretty obvious that it does apply to H2 as well. Similarly, http-request may definitely be honored as soon as a HEADERS frame starts to appear while there is no stream. This will also allow to deal with too long CONTINUATION frames. This patch moves the timeout update to a new function, h2c_update_timeout(), which is in charge of this. It also adds an "idle_start" timestamp in the connection, which is set when nb_cs reaches zero or when a headers frame start to arrive, so that it cannot be delayed too long. This patch should be backported to recent stable releases after some observation time. It depends on previous patch "MEDIUM: mux-h2: slightly relax timeout management rules".	2022-03-18 17:43:34 +01:00
Willy Tarreau	3439583dd6	MEDIUM: mux-h2: slightly relax timeout management rules The H2 timeout rules were arranged to cover complex situations In 2.1 with commit `c2ea47fb1` ("BUG/MEDIUM: mux-h2: do not enforce timeout on long connections"). It turns out that such rules while complex, do not perfectly cover all use cases. The real intent is to say that as long as there are attached streams, the connection must not timeout. Then once all these streams have quit (possibly for timeout reasons) then the mux should take over the management of timeouts. We do have this nb_cs field which indicates the number of attached streams, and it's updated even when leaving orphaned streams. So checking it alone is sufficient to know whether it's the mux or the streams that are in charge of the timeouts. In its current state, this doesn't cause visible effects except that it makes it impossible to implement more subtle parsing timeouts. This would need to be backported as far as 2.0 along with the next commit that will depend on it.	2022-03-18 17:43:34 +01:00
Willy Tarreau	6e805dab2a	BUG/MEDIUM: trace: avoid race condition when retrieving session from conn->owner There's a rare race condition possible when trying to retrieve session from a back connection's owner, that was fixed in 2.4 and described in commit `3aab17bd5` ("BUG/MAJOR: connection: reset conn->owner when detaching from session list"). It also affects the trace code which does the same, so the same fix is needed, i.e. check from conn->session_list that the connection is still enlisted. It's visible when sending a few tens to hundreds of parallel requests to an h2 backend and enabling traces in parallel. This should be backported as far as 2.2 which is the oldest version supporting traces.	2022-03-18 17:43:28 +01:00
Willy Tarreau	d1480cc8a4	BUG/MEDIUM: stream-int: do not rely on the connection error once established Historically the stream-interface code used to check for connection errors by itself. Later this was partially deferred to muxes, but only once the mux is installed or the connection is at least in the established state. But probably as a safety practice the connection error tests remained. The problem is that they are causing trouble on when a response received from a mux is mixed with an error report. The typical case is an upload that is interrupted by the server sending an error or redirect without draining all data, causing an RST to be queued just after the data. In this case the mux has the data, the CO_FL_ERROR flag is present on the connection, and unfortunately the stream-interface refuses to retrieve the data due to this flag, and return an error to the client. It's about time to only rely on CS_FL_ERROR which is set by the mux, but the stream-interface is still responsible for the connection during its setup. However everywhere the CO_FL_ERROR is checked, CS_FL_ERROR is also checked. This commit addresses this by: - adding a new function si_is_conn_error() that checks the SI state and only reports the status of CO_FL_ERROR for states before SI_ST_EST. - eliminating all checks for CO_FL_ERORR in places where CS_FL_ERROR is already checked and either the presence of a mux was already validated or the stream-int's state was already checked as being SI_ST_EST or higher. CO_FL_ERROR tests on the send() direction are also inappropriate as they may cause the loss of pending data. Now this doesn't happen anymore and such events are only converted to CS_FL_ERROR by the mux once notified of the problem. As such, this must not cause the loss of any error event. Now an early error reported on a backend mux doesn't prevent the queued response from being read and forwarded to the client (the list of syscalls below was trimmed and epoll_ctl is not represented): recvfrom(10, "POST / HTTP/1.1\r\nConnection: clo"..., 16320, 0, NULL, NULL) = 66 sendto(11, "POST / HTTP/1.1\r\ntransfer-encodi"..., 47, MSG_DONTWAIT\|MSG_NOSIGNAL, NULL, 0) = 47 epoll_wait(3, [{events=EPOLLIN\|EPOLLERR\|EPOLLHUP\|EPOLLRDHUP, data={u32=11, u64=11}}], 200, 15001) = 1 recvfrom(11, "HTTP/1.1 200 OK\r\ncontent-length:"..., 16320, 0, NULL, NULL) = 57 sendto(10, "HTTP/1.1 200 OK\r\ncontent-length:"..., 57, MSG_DONTWAIT\|MSG_NOSIGNAL, NULL, 0) = 57 epoll_wait(3, [{events=EPOLLIN\|EPOLLERR\|EPOLLHUP\|EPOLLRDHUP, data={u32=11, u64=11}}], 200, 13001) = 1 epoll_wait(3, [{events=EPOLLIN, data={u32=10, u64=10}}], 200, 13001) = 1 recvfrom(10, "A\n0123456789\r\n0\r\n\r\n", 16320, 0, NULL, NULL) = 19 shutdown(10, SHUT_WR) = 0 close(11) = 0 close(10) = 0 Above the server is an haproxy configured with the following: listen blah bind :8002 mode http timeout connect 5s timeout client 5s timeout server 5s option httpclose option nolinger http-request return status 200 hdr connection close And the client takes care of sending requests and data in two distinct parts: while :; do ./dev/tcploop/tcploop 8001 C T S:"POST / HTTP/1.1\r\nConnection: close\r\nTransfer-encoding: chunked\r\n\r\n" P1 S:"A\n0123456789\r\n0\r\n\r\n" P R F; done With this, a small percentage of the requests will reproduce the behavior above. Note that this fix requires the following patch to be applied for the test above to work: BUG/MEDIUM: mux-h1: only turn CO_FL_ERROR to CS_FL_ERROR with empty ibuf This should be backported with after a few weeks of observation, and likely one version at a time. During the backports, the patch might need to be adjusted at each check of CO_FL_ERORR to follow the principles explained above.	2022-03-18 17:00:19 +01:00
Willy Tarreau	99bbdbcc21	BUG/MEDIUM: mux-h1: only turn CO_FL_ERROR to CS_FL_ERROR with empty ibuf A connection-level error must not be turned to a stream-level error if there are still pending data for that stream, otherwise it can cause the truncation of the last pending data. This must be backported to affected releases, at least as far as 2.4, maybe further.	2022-03-18 17:00:17 +01:00
William Lallemand	58a81aeb91	BUG/MINOR: httpclient: CF_SHUTW_NOW should be tested with channel_is_empty() CF_SHUTW_NOW shouldn't be a condition alone to exit the io handler, it must be tested with the emptiness of the response channel. Must be backported to 2.5.	2022-03-18 11:34:10 +01:00
William Lallemand	1eca894321	BUG/MINOR: httpclient: process the response when received before the end of the request A server could reply a response with a shut before the end of the htx transfer, in this case the httpclient would leave before computing the received response. This patch fixes the issue by calling the "process_data" label instead of the "more" label which don't do the si_shut. Must be bacported in 2.5.	2022-03-18 11:34:10 +01:00
William Lallemand	a625b03e83	BUG/MINOR: httpclient: only check co_data() instead of HTTP_MSG_DATA Checking msg >= HTTP_MSG_DATA was useful to check if we received all the data. However it does not work correctly in case of errors because we don't reach this state, preventing to catch the error in the httpclient. The consequence of this problem is that we don't get the status code of the error response upon an error. Fix the issue by only checking co_data(). Must be backported to 2.5.	2022-03-18 11:34:10 +01:00
Willy Tarreau	dd7e6c6dc7	BUG/MINOR: http-rules: completely free incorrect TCP rules on error When a http-request or http-response rule fails to parse, we currently free only the rule without its contents, which makes ASAN complain. Now that we have a new function for this, let's completely free the rule. This relies on this commit: MINOR: actions: add new function free_act_rule() to free a single rule It's probably not needed to backport this since we're on the exit path anyway.	2022-03-17 20:29:06 +01:00
Willy Tarreau	388c0f2a63	BUG/MINOR: tcp-rules: completely free incorrect TCP rules on error When a tcp-request or tcp-response rule fails to parse, we currently free only the rule without its contents, which makes ASAN complain. Now that we have a new function for this, let's completely free the rule. Reg-tests are now completely OK with ASAN. This relies on this commit: MINOR: actions: add new function free_act_rule() to free a single rule It's probably not needed to backport this since we're on the exit path anyway.	2022-03-17 20:26:54 +01:00
Willy Tarreau	6a783e499c	MINOR: actions: add new function free_act_rule() to free a single rule There was free_act_rules() that frees all rules from a head but nothing to free a single rule. Currently some rulesets partially free their own rules on parsing error, and we're seeing some regtests emit errors under ASAN because of this. Let's first extract the code to free a rule into its own function so that it becomes possible to use it on a single rule.	2022-03-17 20:26:19 +01:00
Willy Tarreau	211ea252d9	BUG/MINOR: logs: fix logsrv leaks on clean exit Log servers are a real mess because: - entries are duplicated using memcpy() without their strings being reallocated, which results in these ones not being freeable every time. - a new field, ring_name, was added in 2.2 by commit `99c453df9` ("MEDIUM: ring: new section ring to declare custom ring buffers.") but it's never initialized during copies, causing the same issue - no attempt is made at freeing all that. Of course, running "haproxy -c" under ASAN quickly notices that and dumps a core. This patch adds the missing strdup() and initialization where required, adds a new free_logsrv() function to cleanly free() such a structure, calls it from the proxy when iterating over logsrvs instead of silently leaking their file names and ring names, and adds the same logsrv loop to the proxy_free_defaults() function so that we don't leak defaults sections on exit. It looks a bit entangled, but it comes as a whole because all this stuff is inter-dependent and was missing. It's probably preferable not to backport this in the foreseable future as it may reveal other jokes if some obscure parts continue to memcpy() the logsrv struct.	2022-03-17 19:53:46 +01:00
William Lallemand	43c2ce4d81	BUG/MINOR: server/ssl: free the SNI sample expression ASAN complains about the SNI expression not being free upon an haproxy -c. Indeed the httpclient is now initialized with a sni expression and this one is never free in the server release code. Must be backported in 2.5 and could be backported in every stable versions.	2022-03-16 18:03:15 +01:00
William Lallemand	715c101a19	BUILD: httpclient: fix build without SSL src/http_client.c: In function ‘httpclient_cfg_postparser’: src/http_client.c:1065:8: error: unused variable ‘errmsg’ [-Werror=unused-variable] 1065 \| char *errmsg = NULL; \| ^~~~~~ src/http_client.c:1064:6: error: unused variable ‘err_code’ [-Werror=unused-variable] 1064 \| int err_code = 0; \| ^~~~~~~~ Fix the build of the httpclient without SSL, the problem was introduced with previous patch `71e3158` ("BUG/MINOR: httpclient: send the SNI using the host header") Must be backported in 2.5 as well.	2022-03-16 16:39:23 +01:00
William Lallemand	71e3158395	BUG/MINOR: httpclient: send the SNI using the host header Generate an SNI expression which uses the Host header of the request. This is mandatory for most of the SSL servers nowadays. Must be backported in 2.5 with the previous patch which export server_parse_sni_expr().	2022-03-16 15:55:30 +01:00
William Lallemand	0d05867e78	MINOR: server: export server_parse_sni_expr() function Export the server_parse_sni_expr() function in order to create a SNI expression in a server which was not parsed from the configuration.	2022-03-16 15:55:30 +01:00
Christopher Faulet	53fa787a07	BUG/MEDIUM: sink: Properly get the stream-int in appctx callback functions The appctx owner is not a stream-interface anymore. It is now a conn-stream. However, sink code was not updated accordingly. It is now fixed. It is 2.6-specific, no backport is needed.	2022-03-16 10:01:30 +01:00
Christopher Faulet	fe14af30ec	BUG/MEDIUM: cli/debug: Properly get the stream-int in all debug I/O handlers The appctx owner is not a stream-interface anymore. It is now a conn-stream. In the cli I/O handler for the command "debug dev fd", we still handle it as a stream-interface. It is now fixed. It is 2.6-specific, no backport is needed.	2022-03-16 09:52:13 +01:00
Christopher Faulet	9affa931cd	BUG/MEDIUM: applet: Don't call .release callback function twice Since the CS/SI refactoring, the .release callback function may be called twice. The first call when a shutdown for read or for write is performed. The second one when the applet is detached from its conn-stream. The second call must be guarded, just like the first one, to only be performed is the stream-interface is not the in disconnected (SI_ST_DIS) or closed (SI_ST_CLO) state. To simplify the fix, we now always rely on si_applet_release() function. It is 2.6-specific, no backport is needed.	2022-03-15 11:47:53 +01:00
William Lallemand	8f170c7fca	BUG/MINOR: httpclient/lua: stuck when closing without data The httpclient lua code is lacking the end callback, which means it won't be able to wake up the lua code after a longjmp if the connection was closed without any data. Must be backported to 2.5.	2022-03-15 11:42:38 +01:00
Fr�d�ric L�caille	e9a974a37a	BUG/MAJOR: quic: Possible crash with full congestion control window This commit reverts this one: "d5066dd9d BUG/MEDIUM: quic: qc_prep_app_pkts() retries on qc_build_pkt() failures" After having filled the congestion control window, qc_build_pkt() always fails. Then depending on the relative position of the writer and reader indexes for the TX buffer, this could lead this function to try to reuse the buffer even if not full. In such case, we do not always mark the end of the data in this TX buffer. This is something the reader cannot understand: it reads a false datagram length, then a wrong packet address from the TX buffer, leading to an invalid pointer dereferencing.	2022-03-15 10:38:48 +01:00
Fr�d�ric L�caille	2ee5c8b3dd	BUG/MEDIUM: quic: Blocked STREAM when retransmitted STREAM frames which are not acknowledged in order are inserted in ->tx.acked_frms tree ordered by the STREAM frame offset values. Then, they are consumed in order by qcs_try_to_consume(). But, when we retransmit frames, we possibly have to insert the same STREAM frame node (with the same offset) in this tree. The problem is when they have different lengths. Unfortunately the restransmitted frames are not inserted because of the tree nature (EB_ROOT_UNIQUE). If the STREAM frame which has been successfully inserted has a smaller length than the retransmitted ones, when it is consumed they are tailing bytes in the STREAM (retransmitted ones) which indefinitively remains in the STREAM TX buffer which will never properly be consumed, leading to a blocking state. At this time this may happen because we sometimes build STREAM frames with null lengths. But this is another issue. The solution is to use an EB_ROOT tree to support the insertion of STREAM frames with the same offset but with different lengths. As qcs_try_to_consume() support the STREAM frames retransmission this modification should not have any impact.	2022-03-15 10:38:48 +01:00
William Lallemand	97f69c6fb5	BUG/MEDIUM: httpclient: must manipulate head, not first The httpclient mistakenly use the htx_get_first{_blk}() functions instead of the htx_get_head{_blk}() functions. Which could stop the httpclient because it will be without the start line, waiting for data that won't never come. Must be backported in 2.5.	2022-03-14 15:10:12 +01:00
William Lallemand	c020b2505d	BUG/MINOR: httpclient: remove the UNUSED block when parsing headers Remove the UNUSED blocks when iterating on headers, we should not stop when encountering one. We should only stop iterating once we found the EOH block. It doesn't provoke a problem, since we don't manipulates the headers before treating them, but it could evolve in the future. Must be backported to 2.5.	2022-03-14 15:10:12 +01:00
William Lallemand	c8f1eb99b4	BUG/MINOR: httpclient: consume partly the blocks when necessary Consume partly the blocks in the httpclient I/O handler when there is not enough room in the destination buffer for the whole block or when the block is not contained entirely in the channel's output. It prevents the I/O handler to be stuck in cases when we need to modify the buffer with a filter for exemple. Must be backported in 2.5.	2022-03-14 15:10:12 +01:00
William Lallemand	2b7dc4edb0	BUG/MEDIUM: httpclient: don't consume data before it was analyzed In httpclient_applet_io_handler(), on the response path, we don't check if the data are in the output part of the channel, and could consume them before they were analyzed. To fix this issue, this patch checks for the stline and the headers if the msg_state is >= HTTP_MSG_DATA which means the stline and headers were analyzed. For the data part, it checks if each htx blocks is in the output before copying it. Must be backported in 2.5.	2022-03-14 15:10:12 +01:00
Amaury Denoyelle	76e8b70e43	MEDIUM: server: remove experimental-mode for dynamic servers Dynamic servers feature is now judged to be stable enough. Remove the experimental-mode requirement for "add/del server" commands. This should facilitate dynamic servers adoption.	2022-03-11 14:28:28 +01:00
Amaury Denoyelle	7d098bea2b	MEDIUM: check: do not auto configure SSL/PROXY for dynamic servers For server checks, SSL and PROXY is automatically inherited from the server settings if no specific check port is specified. Change this behavior for dynamic servers : explicit "check-ssl"/"check-send-proxy" are required for them. Without this change, it is impossible to add a dynamic server with SSL/PROXY settings and checks without, if the check port is not explicit. This is because "no-check-ssl"/"no-check-send-proxy" keywords are not available for dynamic servers. This change respects the principle that dynamic servers on the CLI should not reuse the same shortcuts used during the config file parsing. Mostly because we expect this feature to be manipulated by automated tools, contrary to the config file which should aim to be the shortest possible for human readability. Update the documentation of the "check" keyword to reflect this change.	2022-03-11 14:28:28 +01:00
Amaury Denoyelle	6ccfa3c40f	MEDIUM: mux-quic: improve bidir STREAM frames sending The current implementation of STREAM frames emission has some limitation. Most notably when we cannot sent all frames in a single qc_send run. In this case, frames are left in front of the MUX list. It will be re-send individually before other frames, possibly another frame from the same STREAM with new data. An opportunity to merge the frames is lost here. This method is now improved. If a frame cannot be send entirely, it is discarded. On the next qc_send run, we retry to send to this position. A new field qcs.sent_offset is used to remember this. A new frame list is used for each qc_send. The impact of this change is not precisely known. The most notable point is that it is a more logical method of emission. It might also improve performance as we do not keep old STREAM frames which might delay other streams.	2022-03-11 11:37:31 +01:00
Amaury Denoyelle	54445d04e4	MINOR: quic: implement sending confirmation Implement a new MUX function qcc_notify_send. This function must be called by the transport layer to confirm the sending of STREAM data to the MUX. For the moment, the function has no real purpose. However, it will be useful to solve limitations on push frame and implement the flow control.	2022-03-11 11:37:31 +01:00
Amaury Denoyelle	db5d1a1b19	MINOR: mux-quic: improve opportunistic retry sending for STREAM frames For the moment, the transport layer function qc_send_app_pkts lacks features. Most notably, it only send up to a single Tx buffer and won't retry even if there is frames left and its Tx buffer is now empty. To overcome this limitation, the MUX implements an opportunistic retry sending mechanism. qc_send_app_pkts is repeatedly called until the transport layer is blocked on an external condition (such as congestion control or a sendto syscall error). The blocking was detected by inspecting the frame list before and after qc_send_app_pkts. If no frame has been poped by the function, we considered the transport layer to be blocked and we stop to send. The MUX is subscribed on the lower layer to send the frames left. However, in case of STREAM frames, qc_send_app_pkts might use only a portion of the data and update the frame offset. So, for STREAM frames, a new mechanism is implemented : if the offset field of the first frame has not been incremented, it means the transport layer is blocked. This should improve transfers execution. Before this change, there is a possibility of interrupted transfer if the mux has not sent everything possible and is waiting on a transport signaling which will never happen. In the future, qc_send_app_pkts should be extended to retry sending by itself. All this code burden will be removed from the MUX.	2022-03-11 11:37:31 +01:00
Amaury Denoyelle	e2ec9421ea	MINOR: mux-quic: prevent push frame for unidir streams For the moment, unidirectional streams handling is not identical to bidirectional ones in MUX/H3 layer, both in Rx and Tx path. As a safety, skip over uni streams in qc_send. In fact, this change has no impact because qcs.tx.buf is emptied before we start using qcs_push_frame, which prevents the call to qcs_push_frame. However, this condition will soon change to improve bidir streams emission, so an explicit check on stream type must be done. It is planified to unify uni and bidir streams handling in a future stage. When implemented, the check will be removed.	2022-03-11 11:37:31 +01:00
Frédéric Lécaille	728b30d750	CLEANUP: quic: Comments fix for qc_prep_(app)pkts() functions Fix the comments for these two functions about their returned values.	2022-03-11 11:37:31 +01:00
Frédéric Lécaille	d5066dd9dd	BUG/MEDIUM: quic: qc_prep_app_pkts() retries on qc_build_pkt() failures The "stop_build" label aim is to try to reuse the TX buffer when there is not enough contiguous room to build a packet. It was defined but not used!	2022-03-11 11:37:31 +01:00
Frédéric Lécaille	530601cd84	MEDIUM: quic: Implement the idle timeout feature The aim of the idle timeout is to silently closed the connection after a period of inactivity depending on the "max_idle_timeout" transport parameters advertised by the endpoints. We add a new task to implement this timer. Its expiry is updated each time we received an ack-eliciting packet, and each time we send an ack-eliciting packet if no other such packet was sent since we received the last ack-eliciting packet. Such conditions may be implemented thanks to QUIC_FL_CONN_IDLE_TIMER_RESTARTED_AFTER_READ new flag.	2022-03-11 11:37:30 +01:00
Frédéric Lécaille	676b849d37	BUG/MINOR: quic: Missing check when setting the anti-amplification limit as reached Ensure the peer address is not validated before setting the anti-amplication limit as reached.	2022-03-11 11:37:30 +01:00
Frédéric Lécaille	f293b69521	MEDIUM: quic: Remove the QUIC connection reference counter There is no need to use such a reference counter anymore since the QUIC connections are always handled by the same thread. quic_conn_drop() is removed. Its code is merged into quic_conn_release().	2022-03-11 11:37:30 +01:00
Willy Tarreau	d2985f3cec	BUG/MINOR: session: fix theoretical risk of memleak in session_accept_fd() Andrew Suffield reported in issue #1596 that we've had a bug in session_accept_fd() since 2.4 with commit `1b3c931bf` ("MEDIUM: connections: Introduce a new XPRT method, start().") where an error label is wrong and may cause the leak of the freshly allocated session in case conn_xprt_start() returns < 0. The code was checked there and the only two transport layers available at this point are raw_sock and ssl_sock. The former doesn't provide a ->start() method hence conn_xprt_start() will always return zero. The second does provide such a function, but it may only return <0 if the underlying transport (raw_sock) has such a method and fails, which is thus not the case. So fortunately it is not possible to trigger this leak. The patch above also touched the accept code in quic_sock() which was mostly a plain copy of the session code, but there the move didn't have this impact, and since then it was simplified and the next change moved it to its final destination with the proper error label. This should be backported as far as 2.4 as a long-term safety measure (e.g. if in the future we have a reason for making conn_xprt_start() to start failing), but will not have any positive nor negative effect in the short term.	2022-03-11 07:25:11 +01:00
Willy Tarreau	0657b93385	MINOR: stream: add "last_rule_file" and "last_rule_line" samples These two sample fetch methods report respectively the file name and the line number where was located the last rule that was final. This is aimed at being used on log-format lines to help admins figure what rule in the configuration gave a final verdict, and help understand the condition that led to the action. For example, it's now possible to log the last matched rule by adding this to the log-format: ... lr=%[last_rule_file]:%[last_rule_line] A regtest is provided to test various combinations of final rules, some even on top of each other from different rulesets.	2022-03-10 11:51:34 +01:00
Willy Tarreau	c6dae869ca	MINOR: rules: record the last http/tcp rule that gave a final verdict When a tcp-{request,response} content or http-request/http-response rule delivers a final verdict (deny, accept, redirect etc), the last evaluated one will now be recorded in the stream. The purpose is to permit to log the last one that performed a final action. For now the log is not produced.	2022-03-10 11:51:34 +01:00
Christopher Faulet	fbff854250	BUG/MAJOR: mux-pt: Always destroy the backend connection on detach In TCP, when a conn-stream is detached from a backend connection, the connection must be always closed. It was only performed if an error or a shutdown occurred or if there was no connection owner. But it is a problem, because, since the 2.3, backend connections are always owned by a session. This way it is possible to have idle connections attached to a session instead of a server. But there is no idle connections in TCP. In addition, when a session owns a connection it is responsible to close it when it is released. But it only works for idle connections. And it only works if the session is released. Thus there is the place for bugs here. And indeed, a connection leak may occur if a connection retry is performed because of a timeout. In this case, the underlying connection is still alive and is waiting to be fully established. Thus, when the conn-stream is detached from the connection, the connection is not closed. Because the PT multiplexer is quite simple, there is no timeout at this stage. We depend on the kenerl to be notified and finally close the connection. With an unreachable server, orphan backend connections may be accumulated for a while. It may be perceived as a leak. Because there is no reason to keep such backend connections, we just close it now. Frontend connections are still closed by the session or when an error or a shutdown occurs. This patch should fix the issue #1522. It must be backported as far as 2.0. Note that the 2.2 and 2.0 are not affected by this bug because there is no owner for backend TCP connections. But it is probably a good idea to backport the patch on these versions to avoid any future bugs.	2022-03-09 15:56:00 +01:00
Tim Duesterhus	a6a3279188	CLEANUP: fcgi: Use `istadv()` in `fcgi_strm_send_params` Found manually, while creating the previous commits to turn `struct proxy` members into ists. There is an existing Coccinelle rule to replace this pattern by `istadv()` in `ist.cocci`: @@ struct ist i; expression e; @@ - i.ptr += e; - i.len -= e; + i = istadv(i, e); But apparently it is not smart enough to match ists that are stored in another struct. It would be useful to make the existing rule more generic, so that it might catch similar cases in the future.	2022-03-09 07:51:27 +01:00
Tim Duesterhus	98f05f6a38	CLEANUP: fcgi: Replace memcpy() on ist by istcat() This is a little cleaner, because the length of the resulting string does not need to be calculated manually.	2022-03-09 07:51:27 +01:00
Tim Duesterhus	b4b03779d0	MEDIUM: proxy: Store server_id_hdr_name as a `struct ist` The server_id_hdr_name is already processed as an ist in various locations lets also just store it as such. see `0643b0e7e` ("MINOR: proxy: Make `header_unique_id` a `struct ist`") for a very similar past commit.	2022-03-09 07:51:27 +01:00
Tim Duesterhus	e502c3e793	MINOR: proxy: Store orgto_hdr_name as a `struct ist` The orgto_hdr_name is already processed as an ist in `http_process_request`, lets also just store it as such. see `0643b0e7e` ("MINOR: proxy: Make `header_unique_id` a `struct ist`") for a very similar past commit.	2022-03-09 07:51:27 +01:00
Tim Duesterhus	b50ab8489e	MINOR: proxy: Store fwdfor_hdr_name as a `struct ist` The fwdfor_hdr_name is already processed as an ist in `http_process_request`, lets also just store it as such. see `0643b0e7e` ("MINOR: proxy: Make `header_unique_id` a `struct ist`") for a very similar past commit.	2022-03-09 07:51:27 +01:00
Tim Duesterhus	4b1fcaaee3	MINOR: proxy: Store monitor_uri as a `struct ist` The monitor_uri is already processed as an ist in `http_wait_for_request`, lets also just store it as such. see `0643b0e7e` ("MINOR: proxy: Make `header_unique_id` a `struct ist`") for a very similar past commit.	2022-03-09 07:51:27 +01:00
Christopher Faulet	5ce1299c64	DEBUG: stream: Fix stream trace message to print response buffer state Channels buffer state is displayed in the strem trace messages. However, because of a typo, the request buffer was used instead of the response one. This patch should be backported as far as 2.2.	2022-03-08 18:31:44 +01:00
Christopher Faulet	5001913033	DEBUG: stream: Add the missing descriptions for stream trace events The description for STRM_EV_FLT_ANA and STRM_EV_FLT_ERR was missing. This patch should be backported as far as 2.2.	2022-03-08 18:31:44 +01:00
Christopher Faulet	e8cefacfa9	BUG/MEDIUM: mcli: Properly handle errors and timeouts during reponse processing The response analyzer of the master CLI only handles read errors. So if there is a write error, the session remains stuck because some outgoing data are blocked in the channel and the response analyzer waits everything to be sent. Because the maxconn is set to 10 for the master CLI, it may be unresponsive if this happens to many times. Now read and write errors, timeouts and client aborts are handled. This patch should solve the issue #1512. It must be backported as far as 2.0.	2022-03-08 18:31:44 +01:00
Christopher Faulet	8b1eed16d0	DEBUG: cache: Update underlying buffer when loading HTX message in cache applet In the I/O handler of the cache applet, we must update the underlying buffer when the HTX message is loaded, using htx_from_buf() function instead of htxbuf(). It is important because the applet will update the message by adding new HTX blocks. This way, the state of the underlying buffer remains consistant with the state of the HTX message. It is especially important if HAProxy is compiled with "DEBUG_STRICT=2" mode. Without this patch, channel_add_input() call crashed if the channel was empty at the begining of the I/O handler. Note that it is more a build/debug issue than a bug. But this patch may prevent future bugs. For now it is safe because htx_to_buf() function is systematically called, updating accordingly the underlying buffer. This patch may be backported as far as 2.0.	2022-03-08 18:29:20 +01:00
Christopher Faulet	e9382e0afe	BUG/MEDIUM: stream: Use the front analyzers for new listener-less streams For now, for a stream, request analyzers are set at 2 stages. The first one is when the stream is created. The session's listener analyzers, if any, are set on the request channel. In addition, some HTTP analyzers are set for HTX streams (AN_REQ_WAIT_HTTP and AN_REQ_HTTP_PROCESS_FE). The second one is when the backend is set on the stream. At the stage, request analyzers are updated using the backend settings. It is an issue for client applets because there is no listener attached to the stream. In addtion, it may have no specific/dedicated backend. Thus, several request analyzers are missing. Among others, the HTTP analyzers for HTTP applets. The HTTP client is the only one affected for now. To fix the bug, when a stream is created without a listener, we use the frontend to set the request analyzers. Note that there is no issue with the response channel because its analyzers are set when the server connection is established. This patch may be backported to all stable versions. Because only the HTTP client is affected, it must at least be backported to 2.5. It is related to the issue #1593.	2022-03-08 18:27:47 +01:00
Christopher Faulet	dbf1e88e87	BUG/MINOR: cache: Set conn-stream/channel EOI flags at the end of request This bug is the same than for the HTTP client. See "BUG/MINOR: httpclient: Set conn-stream/channel EOI flags at the end of request" for details. Note that because a filter is always attached to the stream when the cache is used, there is no issue because there is no direct forwarding in this case. Thus the stream analyzers are able to see the HTX_FL_EOM flag on the HTX messge. This patch must be backported as far as 2.0. But only CF_EOI must be set because applets are not attached to a conn-stream on older versions.	2022-03-08 18:24:16 +01:00
Christopher Faulet	3fa5d19d14	BUG/MINOR: stats: Set conn-stream/channel EOI flags at the end of request This bug is the same than for the HTTP client. See "BUG/MINOR: httpclient: Set conn-stream/channel EOI flags at the end of request" for details. This patch must be backported as far as 2.0. But only CF_EOI must be set because applets are not attached to a conn-stream on older versions.	2022-03-08 18:24:16 +01:00
Christopher Faulet	d8d2708cfe	BUG/MINOR: hlua: Set conn-stream/channel EOI flags at the end of request This bug is the same than for the HTTP client. See "BUG/MINOR: httpclient: Set conn-stream/channel EOI flags at the end of request" for details. This patch must be backported as far as 2.0. But only CF_EOI must be set because applets are not attached to a conn-stream on older versions.	2022-03-08 18:24:16 +01:00
Christopher Faulet	3d4332419c	BUG/MINOR: httpclient: Set conn-stream/channel EOI flags at the end of request In HTX, HTX_FL_EOM flag is added on the message to notifiy the end of the message was received. In addition, the producer must set CS_FL_EOI flag on the conn-stream. If it is a mux, the stream-interface is responsible to set CF_EOI flag on the input channel. But, for now, if the producer is an applet, in addition to the conn-stream flag, it must also set the channel one. These flags are used to notify the stream that the message is finished and no more data are expected. It is especially important when the message itself it directly forwarded from one side to the other. Because in this case, the stream has no way to see the HTX_FL_EOM flag on the message. Otherwise, the stream will detect a client or a server abort, depending on the side. For the HTTP client, it is not really easy to diagnose this error because there is also another bug hiding this one. All HTTP request analyzers are not set on the input channel. This will be fixed by another patch. This patch must be backported to 2.5. It is related to the issue #1593.	2022-03-08 16:33:56 +01:00
Marno Krahmer	a690b73fba	MINOR: stats: Add dark mode support for socket rows In commit `e9ed63e548` dark mode support was added to the stats page. The initial commit does not include dark mode color overwrites for the .socket CSS class. This commit colors socket rows the same way as backends that acre active but do not have a health check defined. This fixes an issue where reading information from socket lines became really hard in dark mode due to suboptimal coloring of the cell background and the font in it.	2022-03-08 14:47:23 +01:00
Amaury Denoyelle	20f89cac95	BUG/MEDIUM: quic: do not drop packet on duplicate stream/decoding error Change the return value to success in qc_handle_bidi_strm_frm for two specific cases : * if STREAM frame is an already received offset * if application decoding failed This ensures that the packet is not dropped and properly acknowledged. Previous to this fix, the return code was set to error which prevented the ACK to be generated. The impact of the bug might be noticeable in environment with packet loss and retransmission. Due to haproxy not generating ACK for packets containing STREAM frames with already received offset, the client will probably retransmit them again, which will worsen the network transmission.	2022-03-08 14:36:32 +01:00
William Lallemand	b0dfd099c5	BUG/MINOR: cli: shows correct mode in "show sess" The "show sess" cli command only handles "http" or "tcp" as a fallback mode, replace this by a call to proxy_mode_str() to show all the modes. Could be backported in every maintained versions.	2022-03-08 12:21:36 +01:00
William Lallemand	06715af9e5	BUG/MINOR: add missing modes in proxy_mode_str() Add the missing PR_MODE_SYSLOG and PR_MODE_PEERS in proxy_mode_str(). Could be backported in every maintained versions.	2022-03-08 12:21:36 +01:00
Willy Tarreau	c4e56dc58c	MINOR: pools: add a new global option "no-memory-trimming" Some users with very large numbers of connections have been facing extremely long malloc_trim() calls on reload that managed to trigger the watchdog! That's a bit counter-productive. It's even possible that some implementations are not perfectly reliable or that their trimming time grows quadratically with the memory used. Instead of constantly trying to work around these issues, let's offer an option to disable this mechanism, since nobody had been complaining in the past, and this was only meant to be an improvement. This should be backported to 2.4 where trimming on reload started to appear.	2022-03-08 10:45:03 +01:00
Frédéric Lécaille	9777ead2ed	CLEANUP: quic: Remove window redundant variable from NewReno algorithm state struct We use the window variable which is stored in the path struct.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	0e7c9a7143	MINOR: quic: More precise window update calculation When in congestion avoidance state and when acknowledging an <acked> number bytes we must increase the congestion window by at most one datagram (<path->mtu>) by congestion window. So thanks to this patch we apply a ratio to the current number of acked bytes : <acked> * <path->mtu> / <cwnd>. So, when <cwnd> bytes are acked we precisely increment <cwnd> by <path->mtu>. Furthermore we take into an account the number of remaining acknowledged bytes each time we increment the window by <acked> storing their values in the algorithm struct state (->remain_acked) so that it might be take into an account at the next ACK event.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	5f6783094d	CLEANUP: quic: Remove useless definitions from quic_cc_event struct Since the persistent congestion detection is done out of the congestion controllers, there is no need to pass them information through quic_cc_event struct. We remove its useless members. Also remove qc_cc_loss_event() which is no more used.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	a5ee0ae6a2	MINOR: quic: Persistent congestion detection outside of controllers We establish the persistent congestion out of any congestion controller to improve the algorithms genericity. This path characteristic detection may be implemented regarless of the underlying congestion control algorithm. Send congestion (loss) event using directly quic_cc_event(), so without qc_cc_loss_event() wrapper function around quic_cc_event(). Take the opportunity of this patch to shorten "newest_time_sent" member field of quic_cc_event to "time_sent".	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	83bfca6c71	MINOR: quic: Add a "slow start" callback to congestion controller We want to be able to make the congestion controllers re-enter the slow start state outside of the congestion controllers themselves. So, we add a callback ->slow_start() to do so. Define this callback for NewReno algorithm.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	ba9db40b07	CLEANUP: quic: Remove QUIC path manipulations out of the congestion controller QUIC connection path in flight bytes is a variable which should not be manipulated by the congestion controller. This latter aim is to compute the congestion window. So, we pass it as less as parameters as possible to do so.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	4d3d36b670	BUG/MINOR: quic: Missing recovery start timer reset The recovery start time must be reset after a persistent congestion has been detected.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	05e30ee7d5	MINOR: quic: Retry on qc_build_pkt() failures This is done going to stop_build label when qc_build_pkt() fails because of a lack of buffer room (returns -1).	2022-03-04 17:47:32 +01:00
David Carlier	43a568575f	BUILD: fix kFreeBSD build. kFreeBSD needs to be treated as a distinct target from FreeBSD since the underlying system libc is the GNU one. Thus, relying only on __GLIBC__ no longer suffice. - freebsd-glibc new target, key difference is including crypt.h and linking to libdl like linux. - cpu affinity available but the api is still the FreeBSD's. - enabling auxiliary data access only for Linux. Patch based on preliminary work done by @bigon. closes #1555	2022-03-04 17:19:12 +01:00
Amaury Denoyelle	c055e30176	MEDIUM: mux-quic: implement MAX_STREAMS emission for bidir streams Implement the locally flow-control streams limit for opened bidirectional streams. Add a counter which is used to count the total number of closed streams. If this number is big enough, emit a MAX_STREAMS frame to increase the limit of remotely opened bidirectional streams. This is the first commit to implement QUIC flow-control. A series of patches should follow to complete this. This is required to be able to handle more than 100 client requests. This should help to validate the Multiplexing interop test.	2022-03-04 17:00:12 +01:00
Amaury Denoyelle	e9c4cc13fc	MINOR: mux-quic: retry send opportunistically for remaining frames This commit should fix the possible transfer interruption caused by the previous commit. The MUX always retry to send frames if there is remaining data after a send call on the transport layer. This is useful if the transport layer is not blocked on the sending path. In the future, the transport layer should retry by itself the send operation if no blocking condition exists. The MUX layer will always subscribe to retry later if remaining frames are reported which indicate a blocking on the transport layer.	2022-03-04 17:00:12 +01:00
Amaury Denoyelle	2c71fe58f0	MEDIUM: mux-quic: use direct send transport API for STREAMs Modify the STREAM emission in qc_send. Use the new transport function qc_send_app_pkts to directly send the list of constructed frames. This allows to remove the tasklet wakeup on the quic_conn and should reduce the latency. If not all frames are send after the transport call, subscribe the MUX on the lower layer to be able to retry. Currently there is a bug because the transport layer does not retry to send frames in excess after a successful sendto. This might cause the transfer to be interrupted.	2022-03-04 17:00:12 +01:00
Amaury Denoyelle	0dc40f06d1	MINOR: mux-quic: complete functions to detect stream type Improve the functions used to detect the stream characteristics : uni/bidirectional and local/remote initiated. Most notably, these functions are now designed to work transparently for a MUX in the frontend or backend side. For this, we use the connection to determine the current MUX side. This will be useful if QUIC is implemented on the server side.	2022-03-04 17:00:12 +01:00
Amaury Denoyelle	749cb647b1	MINOR: mux-quic: refactor transport parameters init Since QUIC accept handling has been improved, the MUX is initialized after the handshake completion. Thus its safe to access transport parameters in qc_init via the quic_conn. Remove quic_mux_transport_params_update which was called by the transport for the MUX. This improves the architecture by removing a direct call from the transport to the MUX. The deleted function body is not transfered to qc_init because this part will change heavily in the near future when implementing the flow-control.	2022-03-04 17:00:12 +01:00
Frédéric Lécaille	c2f561ce1e	MINOR: quic: Export qc_send_app_pkts() This is at least to make this function be callable by the mux.	2022-03-04 17:00:12 +01:00
Frédéric Lécaille	edc81469a8	MINOR: quic: Make qc_build_frms() build ack-eliciting frames from a list We want to be able to build ack-eliciting frames to be embedded into QUIC packets from a prebuilt list of ack-eliciting frames. This will be helpful for the mux which would like to send STREAM frames asap after having builts its own prebuilt list. To do so, we only add a parameter as struct list to this function to handle such a prebuilt list.	2022-03-04 17:00:12 +01:00
Frédéric Lécaille	28c7ea3725	MINOR: quic: Send short packet from a frame list We want to be able to send ack-elicting packets from a list of ack-eliciting frames. So, this patch adds such a paramaters to the function responsible of building 1RTT packets. The entry point function is qc_send_app_pkts() which is used with the underlying packet number space TX frame list as parameter.	2022-03-04 17:00:12 +01:00
Frédéric Lécaille	1c5968b275	MINOR: quic: qc_prep_app_pkts() implementation We want to get rid of the code used during the handshake step. qc_prep_app_pkts() aim is to build short packets which are also datagrams. Make quic_conn_app_io_cb() call this new function to prepare short packets.	2022-03-04 17:00:12 +01:00
Amaury Denoyelle	1455113e93	CLEANUP: quic: complete ABORT_NOW with a TODO comment Add a TODO comment to not forget to properly implement error returned by qcs_push_frame.	2022-03-04 16:56:51 +01:00
Willy Tarreau	3dfb7da04b	CLEANUP: tree-wide: remove a few rare non-ASCII chars As reported by Tim in issue #1428, our sources are clean, there are just a few files with a few rare non-ASCII chars for the paragraph symbol, a few typos, or in Fred's name. Given that Fred already uses the non-accentuated form at other places like on the public list, let's uniformize all this and make sure the code displays equally everywhere.	2022-03-04 08:58:32 +01:00
Willy Tarreau	f9eba78fb8	BUG/MEDIUM: pools: fix ha_free() on area in the process of being freed Commit `e81248c0c` ("BUG/MINOR: pool: always align pool_heads to 64 bytes") added a free of the allocated pool in pool_destroy() using ha_free(), but it added a subtle bug by which once the pool is released, setting its address to NULL inside the structure itself cannot work because the area has just been freed. This will need to be backported wherever the patch above is backported.	2022-03-03 18:42:49 +01:00
Amaury Denoyelle	2d0f873cd8	BUG/MINOR: quic: fix segfault on CC if mux uninitialized A segfault happens when receiving a CONNECTION_CLOSE during handshake. This is because the mux is not initialized at this stage but the transport layer dereferences it. Fix this by ensuring that the MUX is initialized before. Thanks to Willy for his help on this one. Welcome in the QUIC-men team !	2022-03-03 18:09:37 +01:00
Willy Tarreau	e81248c0c8	BUG/MINOR: pool: always align pool_heads to 64 bytes This is the pool equivalent of commit `97ea9c49f` ("BUG/MEDIUM: fd: always align fdtab[] to 64 bytes"). After a careful code review, it happens that the pool heads are the other structures allocated with malloc/calloc that claim to be aligned to a size larger than what the allocator can offer. While no issue was reported on them, no memset() is performed and no type is large, this is a problem waiting to happen, so better fix it. In addition, it's relatively easy to do by storing the allocation address inside the pool_head itself and use it at free() time. Finally, threads might benefit from the fact that the caches will really be aligned and that there will be no false sharing. This should be backported to all versions where it applies easily.	2022-03-02 18:22:08 +01:00
William Lallemand	10a37360c8	BUG/MEDIUM: httpclient/lua: infinite appctx loop with POST When POSTing a request with a payload, and reusing the same httpclient lua instance, one could encounter a spinning of the httpclient appctx. Indeed the sent counter is not reset between 2 POSTs and the condition for sending the EOM flag is never met. Must fixed issue #1593. To be backported in 2.5.	2022-03-02 16:32:47 +01:00
Willy Tarreau	06e66c84fc	DEBUG: reduce the footprint of BUG_ON() calls Many inline functions involve some BUG_ON() calls and because of the partial complexity of the functions, they're not inlined anymore (e.g. co_data()). The reason is that the expression instantiates the message, its size, sometimes a counter, then the atomic OR to taint the process, and the back trace. That can be a lot for an inline function and most of it is always the same. This commit modifies this by delegating the common parts to a dedicated function "complain()" that takes care of updating the counter if needed, writing the message and measuring its length, and tainting the process. This way the caller only has to check a condition, pass a pointer to the preset message, and the info about the type (bug or warn) for the tainting, then decide whether to dump or crash. Note that this part could also be moved to the function but resulted in complain() always being at the top of the stack, which didn't seem like an improvement. Thanks to these changes, the BUG_ON() calls do not result in uninlining functions anymore and the overall code size was reduced by 60 to 120 kB depending on the build options.	2022-03-02 16:00:42 +01:00
Willy Tarreau	a631b86523	BUILD: tcpcheck: do not declare tcp_check_keywords_register() inline This one is referenced in initcalls by its pointer, it makes no sense to declare it inline. At best it causes function duplication, at worst it doesn't build on older compilers.	2022-03-02 14:54:44 +01:00
Willy Tarreau	4de2cda104	BUILD: trace: do not declare trace_registre_source() inline This one is referenced in initcalls by its pointer, it makes no sense to declare it inline. At best it causes function duplication, at worst it doesn't build on older compilers.	2022-03-02 14:53:00 +01:00
Willy Tarreau	368479c3fc	BUILD: http_rules: do not declare http_*_keywords_registre() inline The 3 functions http_{req,res,after_res}_keywords_register() are referenced in initcalls by their pointer, it makes no sense to declare them inline. At best it causes function duplication, at worst it doesn't build on older compilers.	2022-03-02 14:50:38 +01:00
Willy Tarreau	d318e4e022	BUILD: connection: do not declare register_mux_proto() inline This one is referenced in initcalls by its pointer, it makes no sense to declare it inline. At best it causes function duplication, at worst it doesn't build on older compilers.	2022-03-02 14:46:45 +01:00
Frédéric Lécaille	bd24208673	MINOR: quic: Assemble QUIC TLS flags at the same level Do not distinguish the direction (TX/RX) when settings TLS secrets flags. There is not such a distinction in the RFC 9001. Assemble them at the same level: at the upper context level.	2022-03-01 16:34:03 +01:00
Frédéric Lécaille	9355d50f73	CLEANUP: quic: Indentation fix in qc_prep_pkts() Non-invasive modification.	2022-03-01 16:22:35 +01:00
Frédéric Lécaille	7d845f15fd	CLEANUP: quic: Useless tests in qc_try_rm_hp() There is no need to test <qel>. Furthermore the packet type has already checked by the caller.	2022-03-01 16:22:35 +01:00
Frédéric Lécaille	51c9065f66	MINOR: quic: Drop the packets of discarded packet number spaces This is required since this previous commit: "MINOR: quic: Post handshake I/O callback switching" If not, such packets remain endlessly in the RX buffer and cannot be parsed by the new I/O callback used after the handshake has been confirmed.	2022-03-01 16:22:35 +01:00
Frédéric Lécaille	00e2400fa6	MINOR: quic: Post handshake I/O callback switching Implement a simple task quic_conn_app_io_cb() to be used after the handshakes have completed.	2022-03-01 16:22:35 +01:00
Frédéric Lécaille	5757b4a50e	MINOR: quic: Ensure PTO timer is not set in the past Wakeup asap the timer task when setting its timer in the past. Take also the opportunity of this patch to make simplify quic_pto_pktns(): calling tick_first() is useless here to compare <lpto> with <tmp_pto>.	2022-03-01 16:22:35 +01:00
Christopher Faulet	10c9c74cd1	CLEANUP: stream: Remove useless tests on conn-stream in stream_dump() Since the recent refactoring on the conn-streams, a stream has always a defined frontend and backend conn-streams. Thus, in stream_dump(), there is no reason to still test if these conn-streams are defined. In addition, still in stream_dump(), get the stream-interfaces using the conn-streams and not the opposite. This patch should fix issue #1589 and #1590.	2022-03-01 15:22:05 +01:00
Amaury Denoyelle	0e3010b1bb	MEDIUM: quic: rearchitecture Rx path for bidirectional STREAM frames Reorganize the Rx path for STREAM frames on bidirectional streams. A new function qcc_recv is implemented on the MUX. It will handle the STREAM frames copy and offset calculation from transport to MUX. Another function named qcc_decode_qcs from the MUX can be called by transport each time new STREAM data has been copied. The architecture is now cleaner with the MUX layer in charge of parsing the STREAM frames offsets. This is required to be able to implement the flow-control on the MUX layer. Note that as a convenience, a STREAM frame is not partially copied to the MUX buffer. This simplify the implementation for the moment but it may change in the future to optimize the STREAM frames handling. For the moment, only bidirectional streams benefit from this change. In the future, it may be extended to unidirectional streams to unify the STREAM frames processing.	2022-03-01 11:07:27 +01:00
Amaury Denoyelle	3c4303998f	BUG/MINOR: quic: support FIN on Rx-buffered STREAM frames FIN flag on a STREAM frame was not detected if the frame was previously buffered on qcs.rx.frms before being handled. To fix this, copy the fin field from the quic_stream instance to quic_rx_strm_frm. This is required to properly notify the FIN flag on qc_treat_rx_strm_frms for the MUX layer. Without this fix, the request channel might be left opened after the last STREAM frame reception if there is out-of-order frames on the Rx path.	2022-03-01 11:07:06 +01:00
Amaury Denoyelle	3bf06093dc	MINOR: mux-quic: define flag for last received frame This flag is set when the STREAM frame with FIN set has been received on a qcs instance. For now, this is only used as a BUG_ON guard to prevent against multiple frames with FIN set. It will also be useful when reorganize the RX path and move some of its code in the mux.	2022-03-01 10:52:31 +01:00
Amaury Denoyelle	f77e3435a9	MINOR: quic: handle partially received buffered stream frame Adjust the function to handle buffered STREAM frames. If the offset of the frame was already fully received, discard the frame. If only partially received, compute the difference and copy only the newly offset. Before this change, a buffered frame representing a fully or partially received offset caused the loop to be interrupted. The frame was preserved, thus preventing frames with greater offset to be handled. This may fix some occurences of stalled transfer on the request channel if there is out-of-order STREAM frames on the Rx path.	2022-03-01 10:52:31 +01:00
Amaury Denoyelle	2d2d030522	MINOR: quic: simplify copy of STREAM frames to RX buffer qc_strm_cpy can be simplified by simply using b_putblk which already handle wrapping of the destination buffer. The function is kept to update the frame length and offset fields.	2022-03-01 10:52:31 +01:00
Amaury Denoyelle	850695ab1f	CLEANUP: adjust indentation in bidir STREAM handling function Fix indentation in qc_handle_bidi_strm_frm in if condition.	2022-03-01 10:52:31 +01:00
Tim Duesterhus	cc8348fbc1	MINOR: queue: Replace if() + abort() with BUG_ON() see `5cd4bbd7a` ("BUG/MAJOR: threads/queue: Fix thread-safety issues on the queues management")	2022-03-01 10:14:56 +01:00
Tim Duesterhus	17e6b737d7	MINOR: connection: Transform safety check in PROXYv2 parsing into BUG_ON() With BUG_ON() being enabled by default it is more useful to use a BUG_ON() instead of an effectively never-taken if, as any incorrect assumptions will become much more visible. see `488ee7fb6` ("BUG/MAJOR: proxy_protocol: Properly validate TLV lengths")	2022-02-28 17:59:28 +01:00
Tim Duesterhus	f09af57df5	CLEANUP: connection: Indicate unreachability to the compiler in conn_recv_proxy Transform the unreachability comment into a call to `my_unreachable()` to allow the compiler from benefitting from it. see `d1b15b6e9` ("MINOR: proxy_protocol: Ingest PP2_TYPE_UNIQUE_ID on incoming connections") see `615f81eb5` ("MINOR: connection: Use a `struct ist` to store proxy_authority")	2022-02-28 17:59:28 +01:00
Christopher Faulet	8bc1759f60	DEBUG: stream-int: Fix BUG_ON used to test appctx in si_applet_ops callbacks `693b23bb1` ("MEDIUM: tree-wide: Use unsafe conn-stream API when it is relevant") introduced a regression in DEBUG_STRICT mode because some BUG_ON conditions were inverted. It should ok now. In addition, ALREADY_CHECKED macro was removed from appctx_wakeup() function because it is useless now.	2022-02-28 17:29:11 +01:00
Christopher Faulet	234a10aa9b	BUG/MEDIUM: htx: Fix a possible null derefs in htx_xfer_blks() In htx_xfer_blks() function, when headers or trailers are partially transferred, we rollback the copy by removing copied blocks. Internally, all blocks between <dstref> and <dstblk> are removed. But if the transfer was stopped because we failed to reserve a block, the variable <dstblk> is NULL. Thus, we must not try to remove it. It is unexpected to call htx_remove_blk() in this case. htx_remove_blk() was updated to test <blk> variable inside the existing BUG_ON(). The block must be defined. For now, this bug may only be encountered when H2 trailers are copied. On H2 headers, the destination buffer is empty. Thus a swap is performed. This patch should fix the issue #1578. It must be backported as far as 2.4.	2022-02-28 17:16:55 +01:00
Christopher Faulet	4ab8438362	BUG/MEDIUM: mux-fcgi: Don't rely on SI src/dst addresses for FCGI health-checks When an HTTP health-check is performed in FCGI, we must not rely on the SI source and destination addresses to set default parameters (REMOTE_ADDR/REMOTE_PORT and SERVER_NAME/SERVER_PORT) because the backend conn-stream is not attached to a stream but to a healt-check. Thus, there is no stream-interface. In addition, there is no client connection because it is an "internal" session. Thus, for now, in this case, there is only the server connection that can be used. So src/dst addresses are retrieved from the server connection when the CS application is a health-check. This patch should solve issue #1572. It must be backported to 2.5. Note than the CS api has changed. Thus, on HAProxy 2.5, we should test the session's origin instead: const struct sockaddr_storage src = (cs_check(fstrm->cs) ? ...); const struct sockaddr_storage dst = (cs_check(fstrm->cs) ? ...);	2022-02-28 17:16:55 +01:00
Christopher Faulet	9936dc6577	REORG: stream-int: Uninline si_sync_recv() and make si_cs_recv() private This way si__recv() and si__sned() API are defined the same way. si_sync_snd/si_sync_recv are both exported and defined in the C file. And si_cs_send/si_cs_recv are private and only used by stream-interface internals.	2022-02-28 17:16:47 +01:00
Christopher Faulet	494162381e	CLEANUP: stream-int: Make si_cs_send() function static This function was not exported and is only used in stream_interface.c. So make it static.	2022-02-28 17:13:36 +01:00
Christopher Faulet	693b23bb10	MEDIUM: tree-wide: Use unsafe conn-stream API when it is relevant The unsafe conn-stream API (__cs_*) is now used when we are sure the good endpoint or application is attached to the conn-stream. This avoids compiler warnings about possible null derefs. It also simplify the code and clear up any ambiguity about manipulated entities.	2022-02-28 17:13:36 +01:00
Willy Tarreau	84240044f0	MINOR: channel: don't use co_set_data() to decrement output The use of co_set_data() should be strictly limited to setting the amount of existing data to be transmitted. It ought not be used to decrement the output after the data have left the buffer, because doing so involves performing incorrect calculations using co_data() that still comprises data that are not in the buffer anymore. Let's use c_rew() for this, which is made exactly for this purpose, i.e. decrement c->output by as much as requested. This is cleaner, faster, and will permit stricter checks.	2022-02-28 16:51:23 +01:00
Willy Tarreau	6d3f1e322e	DEBUG: rename WARN_ON_ONCE() to CHECK_IF() The only reason for warning once is to check if a condition really happens. Let's use a term that better translates the intent, that's important when reading the code.	2022-02-28 11:51:23 +01:00
Amaury Denoyelle	7b4c9d6e8c	MINOR: quic: add a TODO for a memleak frame on ACK consume The quic_frame instance containing the quic_stream must be freed when the corresponding ACK has been received. However when implementing this on qcs_try_to_consume, some data transfers are interrupted and cannot complete (DC test from interop test suite).	2022-02-25 15:06:17 +01:00
Amaury Denoyelle	0c7679dd86	MINOR: quic: liberate the TX stream buffer after ACK processing The sending buffer of each stream is cleared when processing ACKs corresponding to STREAM emitted frames. If the buffer is empty, free it and offer it as with other dynamic buffers usage. This should reduce memory consumption as before an opened stream confiscate a buffer during its whole lifetime even if there is no more data to transmit.	2022-02-25 15:06:17 +01:00
Amaury Denoyelle	642ab06313	MINOR: quic: adjust buffer handling for STREAM transmission Simplify the data manipulation of STREAM frames on TX. Only stream data and len field are used to generate a valid STREAM frames from the buffer. Do not use the offset field, which required that a single buffer instance should be shared for every frames on a single stream.	2022-02-25 15:06:17 +01:00
Willy Tarreau	4e0a8b1224	DEBUG: add a new WARN_ON_ONCE() macro This one will maintain a static counter per call place and will only emit the warning on the first call. It may be used to invite users to report an unexpected event without spamming them with messages.	2022-02-25 11:55:47 +01:00
Willy Tarreau	305cfbde43	DBEUG: add a new WARN_ON() macro This is the same as BUG_ON() except that it never crashes and only emits a warning and a backtrace, inviting users to report the problem. This will be usable for non-fatal issues that should not happen and need to be fixed. This way the BUG_ON() when using DEBUG_STRICT_NOCRASH is effectively an equivalent of WARN_ON().	2022-02-25 11:55:47 +01:00
Willy Tarreau	edd426871f	DEBUG: move the tainted stuff to bug.h for easier inclusion The functions needed to manipulate the "tainted" flags were located in too high a level to be callable from the lower code layers. Let's move them to bug.h.	2022-02-25 11:55:38 +01:00
Willy Tarreau	9b4a0e6bac	BUG/MINOR: debug: fix get_tainted() to properly read an atomic value get_tainted() was using an atomic store from the atomic value to a local one instead of using an atomic load. In practice it has no effect given the relatively rare updates of this field and the fact that it's read only when dumping "show info" output, but better fix it. There's probably no need to backport this.	2022-02-25 11:54:30 +01:00
Willy Tarreau	c72d2c7e5b	BUILD: stream: fix build warning with older compilers GCC 6 was not very good at value propagation and is often mislead about risks of null derefs. Since 2.6-dev commit `13a35e575` ("MAJOR: conn_stream/ stream-int: move the appctx to the conn-stream"), it sees a risk of null- deref in stream_upgrade_from_cs() after checking cs_conn_mux(cs). Let's disguise the result so that it doesn't complain anymore. The output code is exactly the same. The same method could be used to shut warnings at -O1 that affect the same compiler by the way.	2022-02-24 19:43:15 +01:00
Amaury Denoyelle	119965f15e	BUG/MEDIUM: quic: fix received ACK stream calculation Adjust the handling of ACK for STREAM frames. When receiving a ACK, the corresponding frames from the acknowledged packet are retrieved. If a frame is of type STREAM, we compare the frame STREAM offset with the last offset known of the qcs instance. The comparison was incomplete as it did not treat a acked offset smaller than the known offset. Previously, the acked frame was incorrectly buffered in the qcs.tx.acked_frms. On reception of future ACKs, when trying to process the buffered acks via qcs_try_to_consume, the loop is interrupted on the smallest offset different from the qcs known offset : in this case it will be the previous smaller range. This is a real bug as it prevents all buffered ACKs to be processed, eventually filling the qcs sending buffer and cause the transfer to stall. Fix this by properly properly handle smaller acked offset. First check if the offset length is greater than the qcs offset and mark as acknowledged the difference on the qcs. If not, the frame is not buffered and simply ignored.	2022-02-24 18:37:39 +01:00
Willy Tarreau	282b6a7539	BUG/MINOR: proxy: preset the error message pointer to NULL in parse_new_proxy() As reported by Coverity in issue #1568, a missing initialization of the error message pointer in parse_new_proxy() may result in displaying garbage or crashing in case of memory allocation error when trying to create a new proxy on startup. This should be backported to 2.4.	2022-02-24 16:40:04 +01:00
Christopher Faulet	2da02ae8b2	BUILD: tree-wide: Avoid warnings about undefined entities retrieved from a CS Since recent changes related to the conn-stream/stream-interface refactoring, GCC reports potential null pointer dereferences when we get the appctx, the stream or the stream-interface from the conn-strem. Of course, depending on the time, these entities may be null. But at many places, we know they are defined and it is safe to get them without any check. Thus, we use ALREADY_CHECKED() macro to silent these warnings. Note that the refactoring is unfinished, so it is not a real issue for now.	2022-02-24 13:56:52 +01:00
Christopher Faulet	9264a2c0e8	BUG/MINOR: h3/hq_interop: Fix CS and stream creation Some recent API changes about conn-stream and stream creation were not fully applied to the H3 part. It is 2.6-DEV specific, no backport is needed.	2022-02-24 11:13:59 +01:00
Christopher Faulet	c983b2114d	CLEANUP: backend: Don't export connect_server anymore connect_server() function is only called from backend.c. So make it static.	2022-02-24 11:00:03 +01:00
Christopher Faulet	54e85cbfc7	MAJOR: check: Use a persistent conn-stream for health-checks In the same way a stream has always valid conn-streams, when a health-checks is created, a conn-stream is now created and the health-check is attached on it, as an app. This simplify a bit the connect part when a health-check is running.	2022-02-24 11:00:03 +01:00
Christopher Faulet	14fd99a20c	MINOR: stream: Don't destroy conn-streams but detach app and endp Don't call cs_destroy() anymore when a stream is released. Instead the endpoint and the app are detached from the conn-stream.	2022-02-24 11:00:03 +01:00
Christopher Faulet	c36de9dc93	MINOR: conn-stream: Release a CS when both app and endp are detached cs_detach_app() function is added to detach an app from a conn-stream. And now, both cs_detach_app() and cs_detach_endp() release the conn-stream when both the app and the endpoint are detached.	2022-02-24 11:00:03 +01:00
Christopher Faulet	014ac35eb2	CLEANUP: stream-int: rename si_reset() to si_init() si_reset() function is only used when a stream-interface is allocated. Thus rename it to si_init() insteaad.	2022-02-24 11:00:03 +01:00
Christopher Faulet	cda94accb1	MAJOR: stream/conn_stream: Move the stream-interface into the conn-stream Thanks to all previous changes, it is now possible to move the stream-interface into the conn-stream. To do so, some SI functions are removed and their conn-stream counterparts are added. In addition, the conn-stream is now responsible to create and release the stream-interface. While the stream-interfaces were inlined in the stream structure, there is now a pointer in the conn-stream. stream-interfaces are now dynamically allocated. Thus a dedicated pool is added. It is a temporary change because, at the end, the stream-interface structure will most probably disappear.	2022-02-24 11:00:03 +01:00
Christopher Faulet	108ce5a70b	MINOR: sink: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the sink part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	0de82720e7	MINOR: tcp-act: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the tcp-act part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	b91afea91c	MINOR: httpclient: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the httpclient part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	e1ede302c3	MINOR: http-act: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the http-act part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	8f8f35b2b0	MINOR: dns: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the dns part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	7a58d79dd2	MINOR: cache: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the cache part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	436811f4a8	MINOR: hlua: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the hlua part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	5d3c8aa154	MINOR: debug: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the debug part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	56489e2e31	MINOR: peers: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the peers part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	4d056bcb70	MINOR: proxy: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the proxy part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	503d26428d	MINOR: frontend: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the frontend part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	02fc86e8f6	MINOR: log: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the log part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	0c247df38b	MINOR: cli: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the cli part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	a629447d02	MINOR: http-ana: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the http-ana part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	5c8b47f665	MINOR: stream: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the stream part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	4a0114b298	MINOR: backend: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the backend part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	0dd566b42e	MINOR: stream: Slightly rework stream_new to separate CS/SI initialization It is just a minor reforctoring of stream_new() function to ease next changes. Especially to move the SI from the stream to the conn-stream.	2022-02-24 11:00:02 +01:00
Christopher Faulet	95a61e8a0e	MINOR: stream: Add pointer to front/back conn-streams into stream struct frontend and backend conn-streams are now directly accesible from the stream. This way, and with some other changes, it will be possible to remove the stream-interfaces from the stream structure.	2022-02-24 11:00:02 +01:00
Christopher Faulet	f835dea939	MEDIUM: conn_stream: Add a pointer to the app object into the conn-stream In the same way the conn-stream has a pointer to the stream endpoint , this patch adds a pointer to the application entity in the conn-stream structure. For now, it is a stream or a health-check. It is mandatory to merge the stream-interface with the conn-stream.	2022-02-24 11:00:02 +01:00
Christopher Faulet	86e1c3381b	MEDIUM: applet: Set the conn-stream as appctx owner instead of the stream-int Because appctx is now an endpoint of the conn-stream, there is no reason to still have the stream-interface as appctx owner. Thus, the conn-stream is now the appctx owner.	2022-02-24 11:00:02 +01:00
Christopher Faulet	13a35e5752	MAJOR: conn_stream/stream-int: move the appctx to the conn-stream Thanks to previous changes, it is now possible to set an appctx as endpoint for a conn-stream. This means the appctx is no longer linked to the stream-interface but to the conn-stream. Thus, a pointer to the conn-stream is explicitly stored in the stream-interface. The endpoint (connection or appctx) can be retrieved via the conn-stream.	2022-02-24 11:00:02 +01:00
Christopher Faulet	dd2d0d8b80	MEDIUM: conn-stream: Be prepared to use an appctx as conn-stream endpoint To be able to use an appctx as conn-stream endpoint, the connection is no longer stored as is in the conn-stream. The obj-type is used instead.	2022-02-24 11:00:02 +01:00
Christopher Faulet	897d612d68	MEDIUM: conn-stream: No longer access connection field directly To be able to handle applets as a conn-stream endpoint, we must be prepared to handle different types of endpoints. First of all, the conn-strream's connection must no longer be used directly.	2022-02-24 11:00:02 +01:00
Christopher Faulet	1329f2a12a	REORG: conn_stream: move conn-stream stuff in dedicated files Move code dealing with the conn-streams in dedicated files.	2022-02-24 11:00:02 +01:00
Christopher Faulet	e2b38b31bb	MEDIUM: stream: Allocate backend CS when the stream is created Because the backend conn-stream is no longer released during connection retry and because it is valid to have conn-stream with no connection, it is possible to allocated it when the stream is created. This means, from now, a stream has always valid frontend and backend conn-streams. It is the first step to merge the SI and the CS.	2022-02-24 11:00:02 +01:00
Christopher Faulet	e00ad358c9	MEDIUM: stream: No longer release backend conn-stream on connection retry The backend conn-stream is no longer released on connection retry. This means the conn-stream is detached from the underlying connection but not released. Thus, during connection retries, the stream has always an allocated conn-stream with no connection. All previous changes were made to make this possible. Note that .attach() mux callback function was changed to get the conn-stream as argument. The muxes are no longer responsible to create the conn-stream when a server connection is attached to a stream.	2022-02-24 11:00:02 +01:00
Christopher Faulet	a742293ec9	MINOR: stream: Handle appctx case first when creating a new stream In the same way the previous commit, when a stream is created, the appctx case is now handled before the conn-stream one. The purpose of this change is to limit bugs during the SI/CS refactoring.	2022-02-24 11:00:01 +01:00
Christopher Faulet	0256da14a5	MINOR: connection: Be prepared to handle conn-stream with no connection The conn-stream will progressively replace the stream-interface. Thus, a stream will have to allocate the backend conn-stream during its creation. This means it will be possible to have a conn-stream with no connection. To prepare this change, we test the conn-stream's connection when we retrieve it.	2022-02-24 11:00:01 +01:00
Willy Tarreau	f4b79c4a01	MINOR: pools: support setting debugging options using -dM The 9 currently available debugging options may now be checked, set, or cleared using -dM. The directive now takes a comma-delimited list of options after the optional poisonning byte. With "help", the list of available options is displayed with a short help and their current status. The management doc was updated.	2022-02-23 17:28:41 +01:00
Willy Tarreau	1408b1f8be	MINOR: pools: delegate parsing of command line option -dM to a new function New function pool_parse_debugging() is now dedicated to parsing options of -dM. For now it only handles the optional memory poisonning byte, but the function may already return an informative message to be printed for help, a warning or an error. This way we'll reuse it for the settings that will be needed for configurable debugging options.	2022-02-23 17:28:41 +01:00
Willy Tarreau	18f96d02d3	MEDIUM: init: handle arguments earlier The argument parser runs too late, we'll soon need it before creating pools, hence just after init_early(). No visible change is expected but this part is sensitive enough to be placed into its own commit for easier bisection later if needed.	2022-02-23 17:28:41 +01:00
Willy Tarreau	392524d222	MINOR: init: extract args parsing to their own function The cmdline argument parsing was performed quite late, which prevents from retrieving elements that can be used to initialize the pools and certain sensitive areas. The goal is to improve this by parsing command line arguments right after the early init stage. This is possible because the cmdline parser already does very little beyond retrieving config elements that are used later. Doing so requires to move the parser code to a separate function and to externalize a few variables out of the function as they're used later in the boot process, in the original function. This patch creates init_args() but doesn't move it upfront yet, it's still executed just before init(), which essentially corresponds to what was done before (only the trash buffers, ACLs and Lua were initialized earlier and are not needed for this). The rest is not modified and as expected no change is observed. Note that the diff doesn't to justice to the change as it makes it look like the early init() code was moved to a new function after the function was renamed, while in fact it's clearly the parser itself which moved.	2022-02-23 17:11:33 +01:00
Willy Tarreau	34527d5354	MEDIUM: init: split the early initialization in its own function There are some delicate chicken-and-egg situations in the initialization code, because the init() function currently does way too much (it goes as far as parsing the config) and due to this it must be started very late. But it's also in charge of initializing a number of variables that are needed in early boot (e.g. hostname/pid for error reporting, or entropy for random generators). This patch carefully extracts all the early code that depends on absolutely nothing, and places it immediately after the STG_LOCK init stage. The only possible failures at this stage are only allocation errors and they continue to provoke an immediate exit(). Some environment variables, hostname, date, pid etc are retrieved at this stage. The program's arguments are also copied there since they're needed to be kept intact for the master process.	2022-02-23 17:11:33 +01:00
Willy Tarreau	3ebe4d989c	MEDIUM: initcall: move STG_REGISTER earlier The STG_REGISTER init level is used to register known keywords and protocol stacks. It must be called earlier because some of the init code already relies on it to be known. For example, "haproxy -vv" for now is constrained to start very late only because of this. This patch moves it between STG_LOCK and STG_ALLOC, which is fine as it's used for static registration.	2022-02-23 17:11:33 +01:00
Willy Tarreau	ef301b7556	MINOR: pools: add a debugging flag for memory poisonning option Now -dM will set POOL_DBG_POISON for consistency with the rest of the pool debugging options. As such now we only check for the new flag, which allows the default value to be preset.	2022-02-23 17:11:33 +01:00
Willy Tarreau	13d7775b06	MINOR: pools: replace DEBUG_MEMORY_POOLS with runtime POOL_DBG_TAG This option used to allow to store a marker at the end of the area, which was used as a canary and detection against wrong freeing while the object is used, and as a pointer to the last pool_free() caller when back in cache. Now that we can compute the offsets at runtime, let's check it at run time and continue the code simplification.	2022-02-23 17:11:33 +01:00
Willy Tarreau	0271822f17	MINOR: pools: replace DEBUG_POOL_TRACING with runtime POOL_DBG_CALLER This option used to allow to store a pointer to the caller of the last pool_alloc() or pool_free() at the end of the area. Now that we can compute the offsets at runtime, let's check it at run time and continue the code simplification. In __pool_alloc() we now always calculate the return address (which is quite cheap), and the POOL_DEBUG_TRACE_CALLER() calls are conditionned on the status of debugging option.	2022-02-23 17:11:33 +01:00
Willy Tarreau	42705d06b7	MINOR: pools: get rid of POOL_EXTRA This macro is build-time dependent and is almost unused, yet where it cannot easily be avoided. Now that we store the distinction between pool->size and pool->alloc_sz, we don't need to maintain it and we can instead compute it on the fly when creating a pool. This is what this patch does. The variables are for now pretty static, but this is sufficient to kill the macro and will allow to set them more dynamically.	2022-02-23 17:11:33 +01:00
Willy Tarreau	96d5bc7379	MINOR: pools: store the allocated size for each pool The allocated size is the visible size plus the extra storage. Since for now we can store up to two extra elements (mark and tracer), it's convenient because now we know that the mark is always stored at ->size, and the tracer is always before ->alloc_sz.	2022-02-23 17:11:33 +01:00
Willy Tarreau	e981631d27	MEDIUM: pools: replace CONFIG_HAP_POOLS with a runtime "NO_CACHE" flag. Like previous patches, this replaces the build-time code paths that were conditionned by CONFIG_HAP_POOLS with runtime paths conditionned by !POOL_DBG_NO_CACHE. One trivial test had to be added in the hot path in __pool_alloc() to refrain from calling pool_get_from_cache(), and another one in __pool_free() to avoid calling pool_put_to_cache(). All cache-specific functions were instrumented with a BUG_ON() to make sure we never call them with cache disabled. Additionally the cache[] array was not initialized (remains NULL) so that we can later drop it if not needed. It's particularly huge and should be turned to dynamic with a pointer to a per-thread area where all the objects are located. This will solve the memory usage issue and will improve locality, or even help better deal with NUMA machines once each thread uses its own arena.	2022-02-23 17:11:33 +01:00
Willy Tarreau	dff3b0627d	MINOR: pools: make the global pools a runtime option. There were very few functions left that were specific to global pools, and even the checks they used to participate to are not directly on the most critical path so they can suffer an extra "if". What's done now is that pool_releasable() always returns 0 when global pools are disabled (like the one before) so that pool_evict_last_items() never tries to place evicted objects there. As such there will never be any object in the free list. However pool_refill_local_from_shared() is bypassed when global pools are disabled so that we even avoid the atomic loads from this function. The default global setting is still adjusted based on the original CONFIG_NO_GLOBAL_POOLS that is set depending on threads and the allocator. The global executable only grew by 1.1kB by keeping this code enabled, and the code is simplified and will later support runtime options.	2022-02-23 17:11:33 +01:00
Willy Tarreau	6f3c7f6e6a	MINOR: pools: add a new debugging flag POOL_DBG_INTEGRITY The test to decide whether or not to enforce integrity checks on cached objects is now enabled at runtime and conditionned by this new debugging flag. While previously it was not a concern to inflate the code size by keeping the two functions static, they were moved to pool.c to limit the impact. In pool_get_from_cache(), the fast code path remains fast by having both flags tested at once to open a slower branch when either POOL_DBG_COLD_FIRST or POOL_DBG_INTEGRITY are set.	2022-02-23 17:11:33 +01:00
Willy Tarreau	d3470e1ce8	MINOR: pools: add a new debugging flag POOL_DBG_COLD_FIRST When enabling pools integrity checks, we usually prefer to allocate cold objects first in order to maximize the time the objects spend in the cache. In order to make this configurable at runtime, let's introduce a new debugging flag to control this allocation order. It is currently preset by the DEBUG_POOL_INTEGRITY build-time setting.	2022-02-23 17:11:33 +01:00
Willy Tarreau	fd8b737e2c	MINOR: pools: switch DEBUG_DONT_SHARE_POOLS to runtime This test used to appear at a single location in create_pool() to enable a check on the pool name or unconditionally merge similarly sized pools. This patch introduces POOL_DBG_DONT_MERGE and conditions the test on this new runtime flag, that is preset according to the aforementioned debugging option.	2022-02-23 17:11:33 +01:00
Willy Tarreau	8d0273ed88	MINOR: pools: switch the fail-alloc test to runtime only The fail-alloc test used to be enabled/disabled at build time using the DEBUG_FAIL_ALLOC macro, but it happens that the cost of the test is quite cheap and that it can be enabled as one of the pool_debugging options. This patch thus introduces the first POOL_DBG_FAIL_ALLOC option, whose default value depends on DEBUG_FAIL_ALLOC. The mem_should_fail() function is now always built, but it was made static since it's never used outside.	2022-02-23 17:11:33 +01:00
Willy Tarreau	605629b008	MINOR: pools: introduce a new pool_debugging global variable This read-mostly variable will be used at runtime to enable/disable certain pool-debugging features and will be set by the command-line parser. A future option -dP will take a number of debugging features as arguments to configure this variable's contents.	2022-02-23 17:11:33 +01:00
Willy Tarreau	af580f659c	MINOR: pools: disable redundant poisonning on pool_free() The poisonning performed on pool_free() used to help a little bit with use-after-free detection, but usually did more harm than good in that it was never possible to perform post-mortem analysis on released objects once poisonning was enabled on allocation. Now that there is a dedicated DEBUG_POOL_INTEGRITY, let's get rid of this annoyance which is not even documented in the management manual.	2022-02-23 17:11:33 +01:00
Willy Tarreau	b61fccdc3f	CLEANUP: init: remove the ifdef on HAPROXY_MEMMAX It's ugly, let's move it to defaults.h with all other ones and preset it to zero if not defined.	2022-02-23 17:11:33 +01:00
Willy Tarreau	cc0d554e5f	CLEANUP: vars: move the per-process variables initialization to vars.c There's no point keeping the vars_init_head() call in init() when we already have a vars_init() registered at the right time to do that, and it complexifies the boot sequence, so let's move it there.	2022-02-23 17:11:33 +01:00
Willy Tarreau	add4306231	CLEANUP: muxes: do not use a dynamic trash in list_mux_protos() Let's not use a trash there anymore. The function is called at very early boot (for "haproxy -vv"), and the need for a trash prevents the arguments from being parsed earlier. Moreover, the function only uses a FILE* on output with fprintf(), so there's not even any benefit in using chunk_printf() on an intermediary variable, emitting the output directly is both clearer and safer.	2022-02-23 17:11:33 +01:00
Willy Tarreau	5b4b6ca823	CLEANUP: httpclient: initialize the client in stage INIT not REGISTER REGISTER is meant to only assemble static lists, not to initialize code that may depend on some elements possibly initialized at this level. For example the init code currently looks up transport protocols such as XPRT_RAW and XPRT_SSL which ought to be themselves registered from at REGISTER stage, and which currently work only because they're still registered directly from a constructor. INIT is perfectly suited for this level.	2022-02-23 17:11:33 +01:00
William Lallemand	ab90ee80d9	BUG/MINOR: httpclient/lua: missing pop for new timeout parameter The lua timeout server lacks a lua_pop(), breaking the lua stack. No backported needed.	2022-02-23 15:16:08 +01:00
William Lallemand	b4a4ef6a29	MINOR: httpclient/lua: ability to set a server timeout Add the ability to set a "server timeout" on the httpclient with either the httpclient_set_timeout() API or the timeout argument in a request. Issue #1470.	2022-02-23 15:11:11 +01:00
Christopher Faulet	686501cb1c	BUG/MEDIUM: stream: Abort processing if response buffer allocation fails In process_stream(), we force the response buffer allocation before any processing to be able to return an error message. It is important because, when an error is triggered, the stream is immediately closed. Thus we cannot wait for the response buffer allocation. When the allocation fails, the stream analysis is stopped and the expiration date of the stream's task is updated before exiting process_stream(). But if the stream was woken up because of a connection or an analysis timeout, the expiration date remains blocked in the past. This means the stream is woken up in loop as long as the response buffer is not properly allocated. Alone, this behavior is already a bug. But because the mechanism to handle buffer allocation failures is totally broken since a while, this bug becomes more problematic. Because, most of time, the watchdog will kill HAProxy in this case because it will detect a spinning loop. To fix it, at least temporarily, an allocation failure at this stage is now reported as an error and the processing is aborted. It's not satisfying but it is better than nothing. If the buffers allocation mechanism is refactored, this part will be reviewed. This patch must be backported, probably as far as 2.0. It may be perceived as a regression, but the actual behavior is probably even worse. And because it was not reported, it is probably not a common situation.	2022-02-23 09:26:32 +01:00
Willy Tarreau	9f699958dc	MINOR: pools: mark most static pool configuration variables as read-mostly The mem_poison_byte, mem_fail_rate, using_default_allocator and the pools list are all only set once at boot time and never changed later, while they're heavily used at run time. Let's optimize their usage from all threads by marking them read-mostly so that them reside in a shared cache line.	2022-02-21 20:44:26 +01:00
Amaury Denoyelle	4323567650	MINOR: quic: fix handling of out-of-order received STREAM frames The recent changes was not complete. `d1c76f24fd` MINOR: quic: do not modify offset node if quic_rx_strm_frm in tree The frame length and data pointer should incremented after the data copy. A BUG_ON statement has been added to detect an incorrect decrement operaiton.	2022-02-21 19:14:09 +01:00
Amaury Denoyelle	c0b66ca73c	MINOR: mux-quic: fix uninitialized return on qc_send This should fix the github issue #1562.	2022-02-21 18:46:58 +01:00
Amaury Denoyelle	ff191de1ca	MINOR: h3: fix compiler warning variable set but not used Some variables were only checked via BUG_ON macro. If compiling without DEBUG_STRICT, this instruction is a noop. Fix this by using an explicit condition + ABORT_NOW. This should fix the github issue #1549.	2022-02-21 18:46:58 +01:00
Amaury Denoyelle	d1c76f24fd	MINOR: quic: do not modify offset node if quic_rx_strm_frm in tree qc_rx_strm_frm_cpy is unsafe because it updates the offset field of the frame. This is not safe as the frame is inserted in the tree when calling this function and offset serves as the key node. To fix this, the API is modified so that qc_rx_strm_frm_cpy does not update the frame parameter. The caller is responsible to update offset/length in case of a partial copy. The impact of this bug is not known. It can only happened with received STREAM frames out-of-order. This might be triggered with large h3 POST requests.	2022-02-21 18:46:58 +01:00
Christopher Faulet	ae17925b87	DEBUG: stream-int: Check CS_FL_WANT_ROOM is not set with an empty input buffer In si_cs_recv(), the mux must never set CS_FL_WANT_ROOM flag on the conn-stream if the input buffer is empty and nothing was copied. It is important because, there is nothing the app layer can do in this case to make some room. If this happens, this will most probably lead to a ping-pong loop between the mux and the stream. With this BUG_ON(), it will be easier to spot such bugs.	2022-02-21 16:29:00 +01:00
Christopher Faulet	ec361bbd84	BUG/MAJOR: mux-h2: Be sure to always report HTX parsing error to the app layer If a parsing error is detected and the corresponding HTX flag is set (HTX_FL_PARSING_ERROR), we must be sure to always report it to the app layer. It is especially important when the error occurs during the response parsing, on the server side. In this case, the RX buffer contains an empty HTX message to carry the flag. And it remains in this state till the info is reported to the app layer. This must be done otherwise, on the conn-stream, the CS_FL_ERR_PENDING flag cannot be switched to CS_FL_ERROR and the CS_FL_WANT_ROOM flag is always set when h2_rcv_buf() is called. The result is a ping-pong loop between the mux and the stream. Note that this patch fixes a bug. But it also reveals a design issue. The error must not be reported at the HTX level. The error is already carried by the conn-stream. There is no reason to duplicate it. In addition, it is errorprone to have an empty HTX message only to report the error to the app layer. This patch should fix the issue #1561. It must be backported as far as 2.0 but the bug only affects HAProxy >= 2.4.	2022-02-21 16:05:47 +01:00
Christopher Faulet	c17c31c822	BUG/MEDIUM: mux-h1: Don't wake h1s if mux is blocked on lack of output buffer After sending some data, we try to wake the H1 stream to resume data processing at the stream level, except if the output buffer is still full. However we must also be sure the mux is not blocked because of an allocation failure on this buffer. Otherwise, it may lead to a ping-pong loop between the stream and the mux to send more data with an unallocated output buffer. Note there is a mechanism to queue buffers allocations when a failure happens. However this mechanism is totally broken since the filters were introducted in HAProxy 1.7. And it is worse now with the multiplexers. So this patch fixes a possible loop needlessly consuming all the CPU. But buffer allocation failures must remain pretty rare. This patch must be backported as far as 2.0.	2022-02-21 16:05:47 +01:00
Amaury Denoyelle	ea3e0355da	MINOR: mux-quic: fix a possible null dereference in qc_timeout_task The qcc instance should be tested as it is implied by a previous test that it may be NULL. In this case, qc_timeout_task can be stopped. This should fix github issue #1559.	2022-02-21 10:05:16 +01:00
Willy Tarreau	11adb1d8fc	BUG/MEDIUM: httpclient: limit transfers to the maximum available room A bug was uncovered by commit `fc5912914` ("MINOR: httpclient: Don't limit data transfer to 1024 bytes"), it happens that callers of b_xfer() and b_force_xfer() are expected to check for available room in the target buffer. Previously it was unlikely to be full but now with full buffer- sized transfers, it happens more often and in practice it is possible to crash the process with the debug command "httpclient" on the CLI by going beyond a the max buffer size. Other call places ought to be rechecked by now and it might be time to rethink this API if it tends to generalize. This must be backported to 2.5.	2022-02-18 17:32:12 +01:00
William Lallemand	8a91374487	BUG/MINOR: tools: url2sa reads ipv4 too far The url2sa implementation is inconsitent when parsing an IPv4, indeed url2sa() takes a <ulen> as a parameter where the call to url2ipv4() takes a null terminated string. Which means url2ipv4 could try to read more that it is supposed to. This function is only used from a buffer so it never reach a unallocated space. It can only cause an issue when used from the httpclient which uses it with an ist. This patch fixes the issue by copying everything in the trash and null-terminated it. Must be backported in all supported version.	2022-02-18 16:32:04 +01:00
Willy Tarreau	2c8f984441	CLEANUP: httpclient/cli: fix indentation alignment of the help message The output was not aligned with other commands, let's fix it.	2022-02-18 16:29:50 +01:00
Remi Tricot-Le Breton	1b01b7f2ef	BUG/MINOR: ssl: Missing return value check in ssl_ocsp_response_print When calling ssl_ocsp_response_print which is used to display an OCSP response's details when calling the "show ssl ocsp-response" on the CLI, we use the BIO_read function that copies an OpenSSL BIO into a trash. The return value was not checked though, which could lead to some crashes since BIO_read can return a negative value in case of error. This patch should be backported to 2.5.	2022-02-18 09:58:04 +01:00
Remi Tricot-Le Breton	8081b67699	BUG/MINOR: ssl: Fix leak in "show ssl ocsp-response" CLI command When calling the "show ssl ocsp-response" CLI command some OpenSSL objects need to be created in order to get some information related to the OCSP response and some of them were not freed. It should be backported to 2.5.	2022-02-18 09:57:57 +01:00
Remi Tricot-Le Breton	a9a591ab3d	BUG/MINOR: ssl: Add missing return value check in ssl_ocsp_response_print The b_istput function called to append the last data block to the end of an OCSP response's detailed output was not checked in ssl_ocsp_response_print. The ssl_ocsp_response_print return value checks were added as well since some of them were missing. This error was raised by Coverity (CID 1469513). This patch fixes GitHub issue #1541. It can be backported to 2.5.	2022-02-18 09:57:51 +01:00
William Lallemand	4f4f2b7b5f	MINOR: httpclient/lua: add 'dst' optionnal field The 'dst' optionnal field on a httpclient request can be used to set an alternative server address in the haproxy address format. Which means it could be use with unix@, ipv6@ etc. Should fix issue #1471.	2022-02-17 20:07:00 +01:00
William Lallemand	7b2e0ee1c1	MINOR: httpclient: sets an alternative destination httpclient_set_dst() allows to set an alternative destination address using HAProxy addres format. This will ignore the address within the URL.	2022-02-17 20:07:00 +01:00
Lukas Tribus	1a16e4ebcb	BUG/MINOR: mailers: negotiate SMTP, not ESMTP As per issue #1552 the mailer code currently breaks on ESMTP multiline responses. Let's negotiate SMTP instead. Should be backported to 2.0.	2022-02-17 15:45:59 +01:00
William Lallemand	5085bc3103	BUG/MINOR: httpclient: reinit flags in httpclient_start() When starting for the 2nd time a request from the same httpclient *hc context, the flags are not reinitialized and the httpclient will stop after the first call to the IO handler, because the END flag is always present. This patch also add a test before httpclient_start() to ensure we don't start a client already started. Must be backported in 2.5.	2022-02-17 12:59:52 +01:00
Willy Tarreau	d0de677682	BUG/MINOR: mux-h2: update the session's idle delay before creating the stream The idle connection delay calculation before a request is a bit tricky, especially for multiplexed protocols. It changed between 2.3 and 2.4 by the integration of the idle delay inside the session itself with these commits: `dd78921c6` ("MINOR: logs: Use session idle duration when no stream is provided") `7a6c51324` ("MINOR: stream: Always get idle duration from the session") and by then it was only set by the H1 mux. But over multiple changes, what used to be a zero idle delay + a request delay for H2 became a bit odd, with the idle time slipping into the request time measurement. The effect is that, as reported in GH issue #1395, some H2 request times look huge. This patch introduces the calculation of the session's idle time on the H2 mux before creating the stream. This is made possible because the stream_new() code immediately copies this value into the stream for use at log time. Thus we don't care about changing something that will be touched by every single request. The idle time is calculated as documented, i.e. the delay from the previous request to the current one. This also means that when a single stream is present on a connection, a part of the server's response time may appear in the %Ti measurement, but this reflects the reality since nothing would prevent the client from using the connection to fetch more objects. In addition this shows how long it takes a client to find references to objects in an HTML page and start to fetch them. A different approach could have consisted in counting from the last time the connection was left without any request (i.e. really idle), but this would at least require a documentation change and it's not certain this would provide a more useful information. Thanks to Bart Butler and Luke Seelenbinder for reporting enough elements to diagnose this issue. This should be backported to 2.4.	2022-02-16 14:42:30 +01:00
Willy Tarreau	c7d85485a0	BUG/MEDIUM: h2/hpack: fix emission of HPACK DTSU after settings change Sadly, despite particular care, commit `39a0a1e12` ("MEDIUM: h2/hpack: emit a Dynamic Table Size Update after settings change") broke H2 when sending DTSU. A missing negation on the flag caused the DTSU_EMITTED flag to be lost and the DTSU to be sent again on the next stream, and possibly to break flow control or a few other internal states. This will have to be backported wherever the patch above was backported. Thanks to Yves Lafon for notifying us with elements to reproduce the issue!	2022-02-16 14:42:13 +01:00
Willy Tarreau	b042e4f6f7	BUG/MAJOR: spoe: properly detach all agents when releasing the applet There's a bug in spoe_release_appctx() which checks the presence of items in the wrong list rt[tid].agents to run over rt[tid].waiting_queue and zero their spoe_appctx. The effect is that these contexts are not zeroed and if spoe_stop_processing() is called, "sa->cur_fpa--" will be applied to one of these recently freed contexts and will corrupt random memory locations, as found at least in bugs #1494 and #1525. This must be backported to all stable versions. Many thanks to Christian Ruppert from Babiel for exchanging so many useful traces over the last two months, testing debugging code and helping set up a similar environment to reproduce it!	2022-02-16 14:42:13 +01:00
Andrew McDermott	bfb15ab34e	BUG/MAJOR: http/htx: prevent unbounded loop in http_manage_server_side_cookies Ensure calls to http_find_header() terminate. If a "Set-Cookie2" header is found then the while(1) loop in http_manage_server_side_cookies() will never terminate, resulting in the watchdog firing and the process terminating via SIGABRT. The while(1) loop becomes unbounded because an unmatched call to http_find_header("Set-Cookie") will leave ctx->blk=NULL. Subsequent calls to check for "Set-Cookie2" will now enumerate from the beginning of all the blocks and will once again match on subsequent passes (assuming a match first time around), hence the loop becoming unbounded. This issue was introduced with HTX and this fix should be backported to all versions supporting HTX. Many thanks to Grant Spence (gspence@redhat.com) for working through this issue with me.	2022-02-16 14:42:13 +01:00
Amaury Denoyelle	1d5fdc526b	MINOR: h3: remove unused return value on decode_qcs This should fix 1470806 coverity report from github issue #1550.	2022-02-16 14:37:56 +01:00
William Lallemand	de6ecc3ace	BUG/MINOR: httpclient/cli: display junk characters in vsn ist are not ended by '\0', leading to junk characters being displayed when using %s for printing the HTTP start line. Fix the issue by replacing %s by %.*s + istlen. Must be backported in 2.5.	2022-02-16 11:37:02 +01:00
Remi Tricot-Le Breton	d544d33e10	BUG/MINOR: jwt: Memory leak if same key is used in multiple jwt_verify calls If the same filename was specified in multiple calls of the jwt_verify converter, we would have parsed the contents of the file every time it was used instead of checking if the entry already existed in the tree. This lead to memory leaks because we would not insert the duplicated entry and we would not free it (as well as the EVP_PKEY it referenced). We now check the return value of ebst_insert and free the current entry if it is a duplicate of an existing entry. The order in which the tree insert and the pkey parsing happen was also switched in order to avoid parsing key files in case of duplicates. Should be backported to 2.5.	2022-02-15 20:08:20 +01:00
Remi Tricot-Le Breton	2b5a655946	BUG/MINOR: jwt: Missing pkey free during cleanup When emptying the jwt_cert_tree during deinit, the entries are freed but not the EVP_PKEY reference they kept, leading in a memory leak. Should be backported in 2.5.	2022-02-15 20:08:20 +01:00
Remi Tricot-Le Breton	4930c6c869	BUG/MINOR: jwt: Double free in deinit function The node pointer was not moving properly along the jwt_cert_tree during the deinit which ended in a double free during cleanup (or when checking a configuration that used the jwt_verify converter with an explicit certificate specified). This patch fixes GitHub issue #1533. It should be backported to 2.5.	2022-02-15 20:08:20 +01:00
Amaury Denoyelle	31e4f6e149	MINOR: h3: report error on HEADERS/DATA parsing Inspect return code of HEADERS/DATA parsing functions and use a BUG_ON to signal an error. The stream should be closed to handle the error in a more clean fashion.	2022-02-15 17:33:21 +01:00
Frédéric Lécaille	71f3abbb52	MINOR: quic: Move quic_rxbuf_pool pool out of xprt part This pool could be confuse with that of the RX buffer pool for the connection (quic_conn_rxbuf).	2022-02-15 17:33:21 +01:00
Frédéric Lécaille	53c7d8db56	MINOR: quic: Do not retransmit too much packets. We retranmist at most one datagram and possibly one more with only PING frame as ack-eliciting frame.	2022-02-15 17:33:21 +01:00
Frédéric Lécaille	0c80e69470	MINOR: quic: Possible frame parsers array overrun This should fix CID 1469663 for GH #1546.	2022-02-15 17:33:21 +01:00
Frédéric Lécaille	59509b5187	MINOR: quic: Non checked returned value for cs_new() in h3_decode_qcs() This should fix CID 1469664 for GH #1546	2022-02-15 17:33:21 +01:00
Frédéric Lécaille	3c08cb4948	MINOR: h3: Dead code in h3_uqs_init() This should fix CID 1469657 for GH #1546.	2022-02-15 17:23:44 +01:00
Frédéric Lécaille	1e1fb5db45	MINOR: quic: Non checked returned value for cs_new() in hq_interop_decode_qcs() This should fix CID 1469657 for GH #1546	2022-02-15 17:23:44 +01:00
Frédéric Lécaille	498e992c1c	MINOR: quic: Useless test in quic_lstnr_dghdlr() This statement is useless. This should fix CID 1469651 for GH #1546.	2022-02-15 17:23:44 +01:00
Frédéric Lécaille	e1c3546efa	MINOR: quic: Avoid warning about NULL pointer dereferences This is the same fixe as for this commit: "BUILD: tree-wide: avoid warnings caused by redundant checks of obj_types" Should fix CID 1469649 for GH #1546	2022-02-15 17:23:44 +01:00
Frédéric Lécaille	ee4508da4f	MINOR: quic: ha_quic_set_encryption_secrets without server specific code Remove this server specific code section. It is useless, not tested. Furthermore this is really not the good place to retrieve the peer transport parameters.	2022-02-15 17:23:44 +01:00
Frédéric Lécaille	16de9f7dbf	MINOR: quic: Code never reached in qc_ssl_sess_init() There was a remaining useless statement in this code block. This fixes CID 1469648 for GH #1546	2022-02-15 17:23:44 +01:00
Frédéric Lécaille	21db6f962b	MINOR: quic: Wrong loss delay computation I really do not know where does this statement come from even after having checked several drafts.	2022-02-15 17:23:44 +01:00
Amaury Denoyelle	91379f79f8	MINOR: h3: implement DATA parsing Add a new function h3_data_to_htx. This function is used to parse a H3 DATA frame and copy it in the mux stream HTX buffer. This is required to support HTTP POST data. Note that partial transfers if the HTX buffer is fulled is not properly handle. This causes large DATA transfer to fail at the moment.	2022-02-15 17:17:00 +01:00
Amaury Denoyelle	7b0f1220d4	MINOR: h3: extract HEADERS parsing in a dedicated function Move the HEADERS parsing code outside of generic h3_decode_qcs to a new dedicated function h3_headers_to_htx. The benefit will be visible when other H3 frames parsing will be implemented such as DATA.	2022-02-15 17:12:27 +01:00
Amaury Denoyelle	0484f92656	MINOR: h3: report frames bigger than rx buffer If a frame is bigger than the qcs buffer, it can not be parsed at the moment. Add a TODO comment to signal that a fix is required.	2022-02-15 17:11:59 +01:00
Amaury Denoyelle	bb56530470	MINOR: h3: set CS_FL_NOT_FIRST When creating a new conn-stream on H3 HEADERS parsing, the flag CS_FL_NOT_FIRST must be set. This is identical to the mux-h2.	2022-02-15 17:10:51 +01:00
Amaury Denoyelle	eb53e5baa1	MINOR: mux-quic: set EOS on rcv_buf Flags EOI/EOS must be set on conn-stream when transfering the last data of a stream in rcv_buf. This is activated if qcs HTX buffer has the EOM flag and has been fully transfered.	2022-02-15 17:10:51 +01:00
Amaury Denoyelle	9a327a7c3f	MINOR: mux-quic: implement rcv_buf Implement the stream rcv_buf operation on QUIC mux. A new buffer is stored in qcs structure named app_buf. This new buffer will contains HTX and will be filled for example on H3 DATA frame parsing. The rcv_buf operation transfer as much as possible data from the HTX from app_buf to the conn-stream buffer. This is mainly identical to mux-h2. This is required to support HTTP POST data.	2022-02-15 17:10:51 +01:00
Amaury Denoyelle	95b93a3a93	MINOR: h3: set properly HTX EOM/BODYLESS on HEADERS parsing Adjust the method to detect that a H3 HEADERS frame is the last one of the stream. If this is true, the flags EOM and BODYLESS must be set on the HTX message.	2022-02-15 17:08:48 +01:00
Amaury Denoyelle	a04724af29	MINOR: h3: add documentation on h3_decode_qcs Specify the purpose of the fin argument on h3_decode_qcs.	2022-02-15 17:08:32 +01:00
Amaury Denoyelle	ffafb3d2c2	MINOR: h3: remove transfer-encoding header According to HTTP/3 specification, transfer-encoding header must not be used in HTTP/3 messages. Remove it when converting HTX responses to HTTP/3.	2022-02-15 17:08:22 +01:00
Amaury Denoyelle	4ac6d37333	BUG/MINOR: h3: fix the header length for QPACK decoding Pass the H3 frame length to QPACK decoding instead of the length of the whole buffer. Without this fix, if there is multiple H3 frames starting with a HEADERS, QPACK decoding will be erroneously applied over all of them, most probably leading to a decoding error.	2022-02-15 17:06:22 +01:00
Amaury Denoyelle	6a2c2f4910	BUG/MINOR: quic: fix FIN stream signaling If the last frame is not entirely copied and must be buffered, FIN must not be signaled to the upper layer. This might fix a rare bug which could cause the request channel to be closed too early leading to an incomplete request.	2022-02-15 17:01:14 +01:00
Amaury Denoyelle	ab9cec7ce1	MINOR: qpack: fix typo in trace hanme -> hname	2022-02-15 11:08:17 +01:00
Amaury Denoyelle	4af6595d41	BUG/MEDIUM: quic: fix crash on CC if mux not present If a CONNECTION_CLOSE is received during handshake or after mux release, a segfault happens due to invalid dereferencement of qc->qcc. Check mux_state first to prevent this.	2022-02-15 11:08:17 +01:00
Amaury Denoyelle	8524f0f779	MINOR: quic: use a global dghlrs for each thread Move the QUIC datagram handlers oustide of the receivers. Use a global handler per-thread which is allocated on post-config. Implement a free function on process deinit to avoid a memory leak.	2022-02-15 10:13:20 +01:00
Willy Tarreau	6c8babf6c4	BUG/MAJOR: sched: prevent rare concurrent wakeup of multi-threaded tasks Since the relaxation of the run-queue locks in 2.0 there has been a very small but existing race between expired tasks and running tasks: a task might be expiring and being woken up at the same time, on different threads. This is protected against via the TASK_QUEUED and TASK_RUNNING flags, but just after the task finishes executing, it releases it TASK_RUNNING bit an only then it may go to task_queue(). This one will do nothing if the task's ->expire field is zero, but if the field turns to zero between this test and the call to __task_queue() then three things may happen: - the task may remain in the WQ until the 24 next days if it's in the future; - the task may prevent any other task after it from expiring during the 24 next days once it's queued - if DEBUG_STRICT is set on 2.4 and above, an abort may happen - since 2.2, if the task got killed in between, then we may even requeue a freed task, causing random behaviour next time it's found there, or possibly corrupting the tree if it gets reinserted later. The peers code is one call path that easily reproduces the case with the ->expire field being reset, because it starts by setting it to TICK_ETERNITY as the first thing when entering the task handler. But other code parts also use multi-threaded tasks and rightfully expect to be able to touch their expire field without causing trouble. No trivial code path was found that would destroy such a shared task at runtime, which already limits the risks. This must be backported to 2.0.	2022-02-14 20:10:43 +01:00
Willy Tarreau	27c8da1fd5	DEBUG: pools: replace the link pointer with the caller's address on pool_free() Along recent evolutions of the pools, we've lost the ability to reliably detect double-frees because while in the past the same pointer was being used to chain the objects in the cache and to store the pool's address, since 2.0 they're different so the pool's address is never overwritten on free() and a double-free will rarely be detected. This patch sets the caller's return address there. It can never be equal to a pool's address and will help guess what was the previous call path. It will not work on exotic architectures nor with very old compilers but these are not the environments where we're trying to get detailed bug reports, and this is not done by default anyway so we don't care about this limitation. Note that depending on the inlining status of the function, the result may differ but that's no big deal either. A test by placing a double free of an appctx inside the release handler itself successfully reported the trouble during appctx_free() and showed that the return address was in stream_int_shutw_applet() (this one calls the release handler).	2022-02-14 20:10:43 +01:00
Willy Tarreau	49bb5d4268	DEBUG: pools: let's add reverse mapping from cache heads to thread and pool During global eviction we're visiting nodes from the LRU tail and we determine their pool cache head and their pool. In order to make sure we never mess up, let's add some backwards pointer to the thread number and pool from the pool_cache_head. It's 64-byte aligned anyway so we're not wasting space and it helps for debugging and will prevent memory corruption the earliest possible.	2022-02-14 20:10:43 +01:00
Willy Tarreau	e2830addda	DEBUG: pools: add extra sanity checks when picking objects from a local cache These few checks are added to make sure we never try to pick an object from an empty list, which would have a devastating effect.	2022-02-14 20:10:43 +01:00
Willy Tarreau	ceabc5ca8c	CLEANUP: pools: don't needlessly set a call mark during refilling of caches When refilling caches from the shared cache, it's pointless to set the pointer to the local pool since it may be overwritten immediately after by the LIST_INSERT(). This is a leftover from the pre-2.4 code in fact. It didn't hurt, though.	2022-02-14 20:10:43 +01:00
Willy Tarreau	c895c441c7	BUG/MINOR: pools: always flush pools about to be destroyed When destroying a pool (e.g. at exit or when resizing buffers), it's important to try to free all their local objects otherwise we can leave some in the cache. This is particularly visible when changing "bufsize", because "show pools" will then show two "trash" pools, one of which contains a single object in cache (which is fortunately not reachable). In all cases this happens while single-threaded so that's easy to do, we just have to do it on the current thread. The easiest way to do this is to pass an extra argument to function pool_evict_from_local_cache() to force a full flush instead of a partial one. This can probably be backported to about all branches where this applies, but at least 2.4 needs it.	2022-02-14 20:10:43 +01:00
Willy Tarreau	b5ba09ed58	BUG/MEDIUM: pools: ensure items are always large enough for the pool_cache_item With the introduction of DEBUG_POOL_TRACING in 2.6-dev with commit `add43fa43` ("DEBUG: pools: add new build option DEBUG_POOL_TRACING"), small pools might be too short to store both the pool_cache_item struct and the caller location, resulting in memory corruption and crashes when this debug option is used. What happens here is that the way the size is calculated is by considering that the POOL_EXTRA part is only used while the object is in use, but this is not true anymore for the caller's pointer which must absolutely be placed after the pool_cache_item. This patch makes sure that the caller part will always start after the pool_cache_item and that the allocation will always be sufficent. This is only tagged medium because the debug option is new and unlikely to be used unless requested by a developer. No backport is needed.	2022-02-14 20:10:43 +01:00
Frédéric Lécaille	547aa0e95e	MINOR: quic: Useless statement in quic_crypto_data_cpy() This should fix Coverity CID 375057 in GH #1526 where a useless assignment was detected.	2022-02-14 15:20:54 +01:00
Frédéric Lécaille	c0b481f87b	MINOR: quic: Possible memleak in qc_new_conn() This should fix Coverity CID 375047 in GH #1536 where <buf_area> could leak because not always freed by by quic_conn_drop(), especially when not stored in <qc> variable.	2022-02-14 15:20:54 +01:00
Frédéric Lécaille	225c31fc9f	CLEANUP: h3: Unreachable target in h3_uqs_init() This should fix Coverity CID 375045 in GH #1536 which detects a no more use "err" target in h3_uqs_init()	2022-02-14 15:20:54 +01:00
Frédéric Lécaille	6842485a84	MINOR: quic: Possible overflow in qpack_get_varint() This should fix CID 375051 in GH 1536 where a signed integer expression (1 << bit) which could overflow was compared to a uint64_t.	2022-02-14 15:20:54 +01:00
Frédéric Lécaille	ce2ecc9643	MINOR: quic: Potential overflow expression in qc_parse_frm() This should fix Coverity CID 375056 where an unsigned char was used to store a 32bit mask.	2022-02-14 15:20:54 +01:00
Frédéric Lécaille	439c464250	MINOR: quic: EINTR error ignored This should fix Coverity CID 375050 in GH #1536 where EINTR errno was ignored due to wrong do...while() loop usage.	2022-02-14 15:20:54 +01:00
Frédéric Lécaille	3916ca197e	MINOR: quic: Variable used before being checked in ha_quic_add_handshake_data() This should fix Coverity CID 375058 in GH issue #1536	2022-02-14 15:20:54 +01:00
Frédéric Lécaille	83cd51e87a	MINOR: quic: Remove an RX buffer useless lock This lock is no more useful: the RX buffer for a connection is always handled by the same thread.	2022-02-14 15:20:54 +01:00
Remi Tricot-Le Breton	88c5695c67	MINOR: ssl: Remove calls to SSL_CTX_set_tmp_dh_callback on OpenSSLv3 The SSL_CTX_set_tmp_dh_callback function was marked as deprecated in OpenSSLv3 so this patch replaces this callback mechanism by a direct set of DH parameters during init.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	c76c3c4e59	MEDIUM: ssl: Replace all DH objects by EVP_PKEY on OpenSSLv3 (via HASSL_DH type) DH structure is a low-level one that should not be used anymore with OpenSSLv3. All functions working on DH were marked as deprecated and this patch replaces the ones we used with new APIs recommended in OpenSSLv3, be it in the migration guide or the multiple new manpages they created. This patch replaces all mentions of the DH type by the HASSL_DH one, which will be replaced by EVP_PKEY with OpenSSLv3 and will remain DH on older versions. It also uses all the newly created helper functions that enable for instance to load DH parameters from a file into an EVP_PKEY, or to set DH parameters into an SSL_CTX for use in a DHE negotiation. The following deprecated functions will effectively disappear when building with OpenSSLv3 : DH_set0_pqg, PEM_read_bio_DHparams, DH_new, DH_free, DH_up_ref, SSL_CTX_set_tmp_dh.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	55d7e782ee	MINOR: ssl: Set default dh size to 2048 Starting from OpenSSLv3, we won't rely on the SSL_CTX_set_tmp_dh_callback mechanism so we will need to know the DH size we want to use during init. In order for the default DH param size to be used when no RSA or DSA private key can be found for a given bind line, we will need to know the default size we want to use (which was not possible the way the code was built, since the global default dh size was set too late.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	bed72631f9	MINOR: ssl: Build local DH of right size when needed The current way the local DH structures are built relies on the fact that the ssl_get_tmp_dh function would only be called as a callback during a DHE negotiation, so after all the SSL contexts are built and the init is over. With OpenSSLv3, this function will now be called during init, so before those objects are curretly built. This patch ensures that when calling ssl_get_tmp_dh and trying to use one of or hard-coded DH parameters, it will be created if it did not exist yet. The current DH parameter creation is also kept so that with versions before OpenSSLv3 we don't end up creating this DH object during a handshake.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	7f6425a130	MINOR: ssl: Add ssl_new_dh_fromdata helper function Starting from OpenSSLv3, the DH_set0_pqg function is deprecated and the use of DH objects directly is advised against so this new helper function will be used to convert our hard-coded DH parameters into an EVP_PKEY. It relies on the new OSSL_PARAM mechanism, as described in the EVP_PKEY-DH manpage.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	5f17930572	MINOR: ssl: Add ssl_sock_set_tmp_dh_from_pkey helper function This helper function will only be used with OpenSSLv3. It simply sets in an SSL_CTX a set of DH parameters of the same size as a certificate's private key. This logic is the same as the one used with older versions, it simply relies on new APIs. If no pkey can be found the SSL_CTX_set_dh_auto function wll be called, making the SSL_CTX rely on DH parameters provided by OpenSSL in case of DHE negotiation.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	846eda91ba	MINOR: ssl: Add ssl_sock_set_tmp_dh helper function Starting from OpenSSLv3, the SSL_CTX_set_tmp_dh function is deprecated and it should be replaced by SSL_CTX_set0_tmp_dh_pkey, which takes an EVP_PKEY instead of a DH parameter. Since this function is new to OpenSSLv3 and its use requires an extra EVP_PKEY_up_ref call, we will keep the two versions side by side, otherwise it would require to get rid of all DH references in older OpenSSL versions as well. This helper function is not used yet so this commit should be strictly iso-functional, regardless of the OpenSSL version.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	292a88ce94	MINOR: ssl: Factorize ssl_get_tmp_dh and append a cbk to its name In the upcoming OpenSSLv3 specific patches, we will make use of the newly created ssl_get_tmp_dh that returns an EVP_PKEY containing DH parameters of the same size as a bind line's RSA or DSA private key. The previously named ssl_get_tmp_dh function was renamed ssl_get_tmp_dh_cbk because it is only used as a callback passed to OpenSSL through SSL_CTX_set_tmp_dh_callback calls.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	09ebb3359a	MINOR: ssl: Add ssl_sock_get_dh_from_bio helper function This new function makes use of the new OpenSSLv3 APIs that should be used to load DH parameters from a file (or a BIO in this case) and that should replace the deprecated PEM_read_bio_DHparams function. Note that this function returns an EVP_PKEY when using OpenSSLv3 since they now advise against using low level structures such as DH ones. This helper function is not used yet so this commit should be stricly iso-functional, regardless of the OpenSSL version.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	78a36e3344	MINOR: ssl: Remove call to ERR_load_SSL_strings with OpenSSLv3 Starting from OpenSSLv3, error strings are loaded automatically so ERR_load_SSL_strings is not needed anymore and was marked as deprecated.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	1effd9aa09	MINOR: ssl: Remove call to ERR_func_error_string with OpenSSLv3 ERR_func_error_string does not return anything anymore with OpenSSLv3, it can be replaced by ERR_peek_error_func which did not exist on previous versions.	2022-02-14 10:07:14 +01:00
William Lallemand	7b820a6191	BUG/MINOR: mworker: does not erase the pidfile upon reload When started in master-worker mode combined with daemon mode, HAProxy will open() with O_TRUNC the pidfile when switching to wait mode. In 2.5, it happens everytime after trying to load the configuration, since we switch to wait mode. In previous version this happens upon a failure of the configuration loading. Fixes bug #1545. Must be backported in every supported branches.	2022-02-14 09:28:13 +01:00
Amaury Denoyelle	58a7704d54	MINOR: quic: take out xprt snd_buf operation Rename quic_conn_to_buf to qc_snd_buf and remove it from xprt ops. This is done to reflect the true usage of this function which is only a wrapper around sendto but cannot be called by the upper layer. qc_snd_buf is moved in quic-sock because to mark its link with quic_sock_fd_iocb which is the recvfrom counterpart.	2022-02-09 15:57:46 +01:00
Amaury Denoyelle	80bd837aaf	MINOR: quic: remove unused xprt rcv_buf operation rcv_buf and the implement quic_conn_to_buf is not used. recvfrom instead is called on the listener polling task via quic_sock_fd_iocb.	2022-02-09 15:57:41 +01:00
Amaury Denoyelle	f6dcbce53e	MINOR: quic: rename local tid variable Rename a local variable tid to cid_tid. This ensures there is no confusion with the global tid. It is now more explicit that we are manipulating a quic datagram handlers from another thread in quic_lstnr_dgram_dispatch.	2022-02-09 15:05:23 +01:00
Amaury Denoyelle	b78805488b	MINOR: h3: hardcode the stream id of control stream Use the value of 0x3 for the stream-id of H3 control-stream. As a consequence, qcs_get_next_id is now unused and is thus removed.	2022-02-09 15:05:23 +01:00
Remi Tricot-Le Breton	c9414e25c4	MINOR: ssl: Remove call to HMAC_Init_ex with OpenSSLv3 HMAC_Init_ex being a function that acts on a low-level HMAC_CTX structure was marked as deprecated in OpenSSLv3. This patch replaces this call by EVP_MAC_CTX_set_params, as advised in the migration_guide, and uses the new OSSL_PARAM mechanism to configure the MAC context, as described in the EVP_MAC and EVP_MAC-HMAC manpages.	2022-02-09 12:11:31 +01:00
Remi Tricot-Le Breton	8ea1f5f6cd	MINOR: ssl: Remove call to SSL_CTX_set_tlsext_ticket_key_cb with OpenSSLv3 SSL_CTX_set_tlsext_ticket_key_cb was deprecated on OpenSSLv3 because it uses an HMAC_pointer which is deprecated as well. According to the v3's manpage it should be replaced by SSL_CTX_set_tlsext_ticket_key_evp_cb which uses a EVP_MAC_CTX pointer. This new callback was introduced in OpenSSLv3 so we need to keep the two calls in the source base and to split the usage depending on the OpenSSL version.	2022-02-09 12:11:31 +01:00
Remi Tricot-Le Breton	c11e7e1d94	MINOR: ssl: Remove EC_KEY related calls when creating a certificate In the context of the 'generate-certificates' bind line option, if an 'ecdhe' option is present on the bind line as well, we use the SSL_CTX_set_tmp_ecdh function which was marked as deprecated in OpenSSLv3. As advised in the SSL_CTX_set_tmp_ecdh manpage, this function should be replaced by the SSL_CTX_set1_groups one (or the SSL_CTX_set1_curves one in our case which does the same but existed on older OpenSSL versions as well). The ECDHE behaviour with OpenSSL 1.0.2 is not the same when using the SSL_CTX_set1_curves function as the one we have on newer versions. Instead of looking for a code that would work exactly the same regardless of the OpenSSL version, we will keep the original code on 1.0.2 and use newer APIs for other versions. This patch should be strictly isofunctional.	2022-02-09 11:15:44 +01:00
Remi Tricot-Le Breton	ff4c3c4c9e	MINOR: ssl: Remove EC_KEY related calls when preparing SSL context The ecdhe option relies on the SSL_CTX_set_tmp_ecdh function which has been marked as deprecated in OpenSSLv3. As advised in the SSL_CTX_set_tmp_ecdh manpage, this function should be replaced by the SSL_CTX_set1_groups one (or the SSL_CTX_set1_curves one in our case which does the same but existed on older OpenSSL versions as well). When using the "curves" option we have a different behaviour with OpenSSL1.0.2 compared to later versions. On this early version an SSL backend using a P-256 ECDSA certificate manages to connect to an SSL frontend having a "curves P-384" option (when it fails with later versions). Even if the API used for later version than OpenSSL 1.0.2 already existed then, for some reason the behaviour is not the same on the older version which explains why the original code with the deprecated API is kept for this version (otherwise we would risk breaking everything on a version that might still be used by some people despite being pretty old). This patch should be strictly isofunctional.	2022-02-09 11:15:44 +01:00
Remi Tricot-Le Breton	2559bc8318	MINOR: ssl: Use high level OpenSSL APIs in sha2 converter The sha2 converter's implementation used low level interfaces such as SHA256_Update which are flagged as deprecated starting from OpenSSLv3. This patch replaces those calls by EVP ones which already existed on older versions. It should be fully isofunctional.	2022-02-09 11:15:44 +01:00
Remi Tricot-Le Breton	36f80f6e0b	CLEANUP: ssl: Remove unused ssl_sock_create_cert function This function is not used anymore, it can be removed.	2022-02-09 11:15:44 +01:00
Remi Tricot-Le Breton	2e7d1eb2a7	BUG/MINOR: ssl: Remove empty lines from "show ssl ocsp-response <id>" output There were empty lines in the output of the CLI's "show ssl ocsp-response <id>" command. The plain "show ssl ocsp-response" command (without parameter) was already managed in commit `cc750efbc5`. This patch adds an extra space to those lines so that the only existing empty lines actually mark the end of the output. This requires to post-process the buffer filled by OpenSSL's OCSP_RESPONSE_print function (which produces the output of the "openssl ocsp -respin <ocsp.pem>" command). This way the output of our command still looks the same as openssl's one. Must be backported in 2.5.	2022-02-03 09:57:24 +01:00
Frédéric Lécaille	bfa3236c6c	MINOR: quic: Remove a useless test in quic_get_dgram_dcid() This test is already done when entering quic_get_dgram_dcid().	2022-02-02 18:24:43 +01:00
Frédéric Lécaille	f6f7520b9b	MINOR: quic: Wrong datagram buffer passed to quic_lstnr_dgram_dispatch() The same datagram could be passed to quic_lstnr_dgram_dispatch() before being consumed by qc_lstnr_pkt_rcv() leading to a wrong decryption for the packet number decryption, then a decryption error for the data. This was due to a wrong datagram buffer passed to quic_lstnr_dgram_dispatch(). The datagram data which must be passed to quic_lstnr_dgram_dispatch() are the same as the one passed to recvfrom().	2022-02-02 18:24:21 +01:00
Frédéric Lécaille	841bf5e7f4	MINOR: quic: Do not modify a marked as consumed datagram Mark the datagrams as consumed at the very last time.	2022-02-02 18:24:21 +01:00
Christopher Faulet	fc5912914b	MINOR: httpclient: Don't limit data transfer to 1024 bytes For debug purpose, no more 1024 bytes were copied at a time. But there is no reason to keep this limitation. Thus, it is removed. This patch may be backported to 2.5.	2022-02-02 16:19:19 +01:00
Christopher Faulet	6ced61dd0a	BUG/MEDIUM: httpclient: Xfer the request when the stream is created Since the HTTP legacy mode was removed, it is unexpected to create an HTTP stream without a valid request. Thanks to this change, the wait_for_request analyzer was significatly simplified. And it is possible because HTTP multiplexers already take care to have a valid request to create a stream. But it means that any HTTP applet on the client side must do the same. The httpclient client is one of them. And it is not a problem because the request is generated before starting the applet. We must just take care to set the right state. For now it works "by chance", because the applet seems to be scheduled before the stream itself. But if this change, this will lead to crash because the stream expects to have a request when wait_for_request analyzer. This patch should be backported to 2.5.	2022-02-02 16:19:17 +01:00
Christopher Faulet	600985df41	BUG/MINOR: httpclient: Revisit HC request and response buffers allocation For now, these buffers are allocated when the httpclient is created and freed when it is released. Usually, we try to avoid to keep buffer allocated if it is not required. Empty buffers should be released ASAP. Apart for that, there is no issue with the response side because a copy is always performed. However, for the request side, a swap with the channel's buffer is always performed. And there is no guarantee the channel's buffer is allocated. Thus, after the swap, the httpclient can retrieve a null buffer. In practice, this never happens. But this may change. And it will be required for a futur fix. So, now, we systematically take care to have an allocated buffer when we want to write in it. And it is released as soon as it becomes empty. This patch should be backported to 2.5.	2022-02-02 16:19:16 +01:00
William Lallemand	dae12c7553	MINOR: mworker/cli: add flags in the prompt The master CLI prompt is now able to show flags in its prompt depending on the mode used: experimental (x), expert (e), mcli-debug (d).	2022-02-02 15:51:24 +01:00
William Lallemand	2a17191e91	MINOR: mworker/cli: mcli-debug-mode enables every command "mcli-debug-mode on" enables every command that were meant for a worker, on the CLI of the master. Which mean you can issue, "show fd", show stat" in order to debug the MASTER proxy. You can also combine it with "expert-mode on" or "experimental-mode on" to access to more commands.	2022-02-02 15:51:24 +01:00
William Lallemand	d9c28070c1	BUG/MINOR: mworker/cli: don't display help on master applet When in expert or experimental mode on the master CLI, and issuing a command for the master process, all commands are prefixed by "mode-experimental -" or/and "mode-expert on -", however these commands were not available in the master applet, so the help was issued for each one.	2022-02-02 15:51:24 +01:00
William Lallemand	fe618fbd0c	CLEANUP: cleanup a commentary in pcli_parse_request() Remove '1' from a commentary in pcli_parse_request()	2022-02-02 15:51:12 +01:00
Willy Tarreau	2454d6ef5b	[RELEASE] Released version 2.6-dev1 Released version 2.6-dev1 with the following main changes : - BUG/MINOR: cache: Fix loop on cache entries in "show cache" - BUG/MINOR: httpclient: allow to replace the host header - BUG/MINOR: lua: don't expose internal proxies - MEDIUM: mworker: seamless reload use the internal sockpairs - BUG/MINOR: lua: remove loop initial declarations - BUG/MINOR: mworker: does not add the -sf in wait mode - BUG/MEDIUM: mworker: FD leak of the eventpoll in wait mode - MINOR: quic: do not reject PADDING followed by other frames - REORG: quic: add comment on rare thread concurrence during CID alloc - CLEANUP: quic: add comments on CID code - MEDIUM: quic: handle CIDs to rattach received packets to connection - MINOR: qpack: support litteral field line with non-huff name - MINOR: quic: activate QUIC traces at compilation - MINOR: quic: use more verbose QUIC traces set at compile-time - MEDIUM: pool: refactor malloc_trim/glibc and jemalloc api addition detections. - MEDIUM: pool: support purging jemalloc arenas in trim_all_pools() - BUG/MINOR: mworker: deinit of thread poller was called when not initialized - BUILD: pools: only detect link-time jemalloc on ELF platforms - CI: github actions: add the output of $CC -dM -E- - BUG/MEDIUM: cli: Properly set stream analyzers to process one command at a time - BUILD: evports: remove a leftover from the dead_fd cleanup - MINOR: quic: Set "no_application_protocol" alert - MINOR: quic: More accurate immediately close. - MINOR: quic: Immediately close if no transport parameters extension found - MINOR: quic: Rename qc_prep_hdshk_pkts() to qc_prep_pkts() - MINOR: quic: Possible crash when inspecting the xprt context - MINOR: quic: Dynamically allocate the secrete keys - MINOR: quic: Add a function to derive the key update secrets - MINOR: quic: Add structures to maintain key phase information - MINOR: quic: Optional header protection key for quic_tls_derive_keys() - MINOR: quic: Add quic_tls_key_update() function for Key Update - MINOR: quic: Enable the Key Update process - MINOR: quic: Delete the ODCIDs asap - BUG/MINOR: vars: Fix the set-var and unset-var converters - MEDIUM: pool: Following up on previous pool trimming update. - BUG/MEDIUM: mux-h1: Fix splicing by properly detecting end of message - BUG/MINOR: mux-h1: Fix splicing for messages with unknown length - MINOR: mux-h1: Improve H1 traces by adding info about http parsers - MINOR: mux-h1: register a stats module - MINOR: mux-h1: add counters instance to h1c - MINOR: mux-h1: count open connections/streams on stats - MINOR: mux-h1: add stat for total count of connections/streams - MINOR: mux-h1: add stat for total amount of bytes received and sent - REGTESTS: h1: Add a script to validate H1 splicing support - BUG/MINOR: server: Don't rely on last default-server to init server SSL context - BUG/MEDIUM: resolvers: Detach query item on response error - MEDIUM: resolvers: No longer store query items in a list into the response - BUG/MAJOR: segfault using multiple log forward sections. - BUG/MEDIUM: h1: Properly reset h1m flags when headers parsing is restarted - BUG/MINOR: resolvers: Don't overwrite the error for invalid query domain name - BUILD: bug: Fix error when compiling with -DDEBUG_STRICT_NOCRASH - BUG/MEDIUM: sample: Fix memory leak in sample_conv_jwt_member_query - DOC: spoe: Clarify use of the event directive in spoe-message section - DOC: config: Specify %Ta is only available in HTTP mode - BUILD: tree-wide: avoid warnings caused by redundant checks of obj_types - IMPORT: slz: use the correct CRC32 instruction when running in 32-bit mode - MINOR: quic: fix segfault on CONNECTION_CLOSE parsing - MINOR: h3: add BUG_ON on control receive function - MEDIUM: xprt-quic: finalize app layer initialization after ALPN nego - MINOR: h3: remove duplicated FIN flag position - MAJOR: mux-quic: implement a simplified mux version - MEDIUM: mux-quic: implement release mux operation - MEDIUM: quic: detect the stream FIN - MINOR: mux-quic: implement subscribe on stream - MEDIUM: mux-quic: subscribe on xprt if remaining data after send - MEDIUM: mux-quic: wake up xprt on data transferred - MEDIUM: mux-quic: handle when sending buffer is full - MINOR: quic: RX buffer full due to wrong CRYPTO data handling - MINOR: quic: Race issue when consuming RX packets buffer - MINOR: quic: QUIC encryption level RX packets race issue - MINOR: quic: Delete remaining RX handshake packets - MINOR: quic: Remove QUIC TX packet length evaluation function - MINOR: hq-interop: fix tx buffering - MINOR: mux-quic: remove uneeded code to check fin on TX - MINOR: quic: add HTX EOM on request end - BUILD: mux-quic: fix compilation with DEBUG_MEM_STATS - MINOR: http-rules: Add capture action to http-after-response ruleset - BUG/MINOR: cli/server: Don't crash when a server is added with a custom id - MINOR: mux-quic: do not release qcs if there is remaining data to send - MINOR: quic: notify the mux on CONNECTION_CLOSE - BUG/MINOR: mux-quic: properly initialize flow control - MINOR: quic: Compilation fix for quic_rx_packet_refinc() - MINOR: h3: fix possible invalid dereference on htx parsing - DOC: config: retry-on list is space-delimited - DOC: config: fix error-log-format example - BUG/MEDIUM: mworker/cli: crash when trying to access an old PID in prompt mode - MINOR: hq-interop: refix tx buffering - REGTESTS: ssl: use X509_V_ERR_UNABLE_TO_GET_ISSUER_CERT_LOCALLY for cert check - MINOR: cli: "show version" displays the current process version - CLEANUP: cfgparse: modify preprocessor guards around numa detection code - MEDIUM: cfgparse: numa detect topology on FreeBSD. - BUILD: ssl: unbreak the build with newer libressl - MINOR: vars: Move UPDATEONLY flag test to vars_set_ifexist - MINOR: vars: Set variable type to ANY upon creation - MINOR: vars: Delay variable content freeing in var_set function - MINOR: vars: Parse optional conditions passed to the set-var converter - MINOR: vars: Parse optional conditions passed to the set-var actions - MEDIUM: vars: Enable optional conditions to set-var converter and actions - DOC: vars: Add documentation about the set-var conditions - REGTESTS: vars: Add new test for conditional set-var - MINOR: quic: Attach timer task to thread for the connection. - CLEANUP: quic_frame: Remove a useless suffix to STOP_SENDING - MINOR: quic: Add traces for STOP_SENDING frame and modify others - CLEANUP: quic: Remove cdata_len from quic_tx_packet struct - MINOR: quic: Enable TLS 0-RTT if needed - MINOR: quic: No TX secret at EARLY_DATA encryption level - MINOR: quic: Add quic_set_app_ops() function - MINOR: ssl_sock: Set the QUIC application from ssl_sock_advertise_alpn_protos. - MINOR: quic: Make xprt support 0-RTT. - MINOR: qpack: Missing check for truncated QPACK fields - CLEANUP: quic: Comment fix for qc_strm_cpy() - MINOR: hq_interop: Stop BUG_ON() truncated streams - MINOR: quic: Do not mix packet number space and connection flags - CLEANUP: quic: Shorten a litte bit the traces in lstnr_rcv_pkt() - MINOR: mux-quic: fix trace on stream creation - CLEANUP: quic: fix spelling mistake in a trace - CLEANUP: quic: rename quic_conn conn to qc in quic_conn_free - MINOR: quic: add missing lock on cid tree - MINOR: quic: rename constant for haproxy CIDs length - MINOR: quic: refactor concat DCID with address for Initial packets - MINOR: quic: compare coalesced packets by DCID - MINOR: quic: refactor DCID lookup - MINOR: quic: simplify the removal from ODCID tree - REGTESTS: vars: Remove useless ssl tunes from conditional set-var test - MINOR: ssl: Remove empty lines from "show ssl ocsp-response" output - MINOR: quic: Increase the RX buffer for each connection - MINOR: quic: Add a function to list remaining RX packets by encryption level - MINOR: quic: Stop emptying the RX buffer asap. - MINOR: quic: Do not expect to receive only one O-RTT packet - MINOR: quic: Do not forget STREAM frames received in disorder - MINOR: quic: Wrong packet refcount handling in qc_pkt_insert() - DOC: fix misspelled keyword "resolve_retries" in resolvers - CLEANUP: quic: rename quic_conn instances to qc - REORG: quic: move mux function outside of xprt - MINOR: quic: add reference to quic_conn in ssl context - MINOR: quic: add const qualifier for traces function - MINOR: trace: add quic_conn argument definition - MINOR: quic: use quic_conn as argument to traces - MINOR: quic: add quic_conn instance in traces for qc_new_conn - MINOR: quic: Add stream IDs to qcs_push_frame() traces - MINOR: quic: unchecked qc_retrieve_conn_from_cid() returned value - MINOR: quic: Wrong dropped packet skipping - MINOR: quic: Handle the cases of overlapping STREAM frames - MINOR: quic: xprt traces fixes - MINOR: quic: Drop asap Retry or Version Negotiation packets - MINOR: pools: work around possibly slow malloc_trim() during gc - DEBUG: ssl: make sure we never change a servername on established connections - MINOR: quic: Add traces for RX frames (flow control related) - MINOR: quic: Add CONNECTION_CLOSE phrase to trace - REORG: quic: remove qc_ prefix on functions which not used it directly - BUG/MINOR: quic: upgrade rdlock to wrlock for ODCID removal - MINOR: quic: remove unnecessary call to free_quic_conn_cids() - MINOR: quic: store ssl_sock_ctx reference into quic_conn - MINOR: quic: remove unnecessary if in qc_pkt_may_rm_hp() - MINOR: quic: replace usage of ssl_sock_ctx by quic_conn - MINOR: quic: delete timer task on quic_close() - MEDIUM: quic: implement refcount for quic_conn - BUG/MINOR: quic: fix potential null dereference - BUG/MINOR: quic: fix potential use of uninit pointer - BUG/MEDIUM: backend: fix possible sockaddr leak on redispatch - BUG/MEDIUM: peers: properly skip conn_cur from incoming messages - CI: Github Actions: do not show VTest failures if build failed - BUILD: opentracing: display warning in case of using OT_USE_VARS at compile time - MINOR: compat: detect support for dl_iterate_phdr() - MINOR: debug: add ability to dump loaded shared libraries - MINOR: debug: add support for -dL to dump library names at boot - BUG/MEDIUM: ssl: initialize correctly ssl w/ default-server - REGTESTS: ssl: fix ssl_default_server.vtc - BUG/MINOR: ssl: free the fields in srv->ssl_ctx - BUG/MEDIUM: ssl: free the ckch instance linked to a server - REGTESTS: ssl: update of a crt with server deletion - BUILD/MINOR: cpuset FreeBSD 14 build fix. - MINOR: pools: always evict oldest objects first in pool_evict_from_local_cache() - DOC: pool: document the purpose of various structures in the code - CLEANUP: pools: do not use the extra pointer to link shared elements - CLEANUP: pools: get rid of the POOL_LINK macro - MINOR: pool: allocate from the shared cache through the local caches - CLEANUP: pools: group list updates in pool_get_from_cache() - MINOR: pool: rely on pool_free_nocache() in pool_put_to_shared_cache() - MINOR: pool: make pool_is_crowded() always true when no shared pools are used - MINOR: pool: check for pool's fullness outside of pool_put_to_shared_cache() - MINOR: pool: introduce pool_item to represent shared pool items - MINOR: pool: add a function to estimate how many may be released at once - MEDIUM: pool: compute the number of evictable entries once per pool - MINOR: pools: prepare pool_item to support chained clusters - MINOR: pools: pass the objects count to pool_put_to_shared_cache() - MEDIUM: pools: centralize cache eviction in a common function - MEDIUM: pools: start to batch eviction from local caches - MEDIUM: pools: release cached objects in batches - OPTIM: pools: reduce local pool cache size to 512kB - CLEANUP: assorted typo fixes in the code and comments This is 29th iteration of typo fixes - CI: github actions: update OpenSSL to 3.0.1 - BUILD/MINOR: tools: solaris build fix on dladdr. - BUG/MINOR: cli: fix _getsocks with musl libc - BUG/MEDIUM: http-ana: Preserve response's FLT_END analyser on L7 retry - MINOR: quic: Wrong traces after rework - MINOR: quic: Add trace about in flight bytes by packet number space - MINOR: quic: Wrong first packet number space computation - MINOR: quic: Wrong packet number space computation for PTO - MINOR: quic: Wrong loss time computation in qc_packet_loss_lookup() - MINOR: quic: Wrong ack_delay compution before calling quic_loss_srtt_update() - MINOR: quic: Remove nb_pto_dgrams quic_conn struct member - MINOR: quic: Wrong packet number space trace in qc_prep_pkts() - MINOR: quic: Useless test in qc_prep_pkts() - MINOR: quic: qc_prep_pkts() code moving - MINOR: quic: Speeding up Handshake Completion - MINOR: quic: Probe Initial packet number space more often - MINOR: quic: Probe several packet number space upon timer expiration - MINOR: quic: Comment fix. - MINOR: quic: Improve qc_prep_pkts() flexibility - MINOR: quic: Do not drop secret key but drop the CRYPTO data - MINOR: quic: Prepare Handshake packets asap after completed handshake - MINOR: quic: Flag asap the connection having reached the anti-amplification limit - MINOR: quic: PTO timer too often reset - MINOR: quic: Re-arm the PTO timer upon datagram receipt - MINOR: proxy: add option idle-close-on-response - MINOR: cpuset: switch to sched_setaffinity for FreeBSD 14 and above. - CI: refactor spelling check - CLEANUP: assorted typo fixes in the code and comments - BUILD: makefile: add -Wno-atomic-alignment to work around clang abusive warning - MINOR: quic: Only one CRYPTO frame by encryption level - MINOR: quic: Missing retransmission from qc_prep_fast_retrans() - MINOR: quic: Non-optimal use of a TX buffer - BUG/MEDIUM: mworker: don't use _getsocks in wait mode - BUG/MINOR: ssl: Store client SNI in SSL context in case of ClientHello error - BUG/MAJOR: mux-h1: Don't decrement .curr_len for unsent data - DOC: internals: document the pools architecture and API - CI: github actions: clean default step conditions - BUILD: cpuset: fix build issue on macos introduced by previous change - MINOR: quic: Remaining TRACEs with connection as firt arg - MINOR: quic: Reset ->conn quic_conn struct member when calling qc_release() - MINOR: quic: Flag the connection as being attached to a listener - MINOR: quic: Wrong CRYPTO frame concatenation - MINOR: quid: Add traces quic_close() and quic_conn_io_cb() - REGTESTS: ssl: Fix ssl_errors regtest with OpenSSL 1.0.2 - MINOR: quic: Do not dereference ->conn quic_conn struct member - MINOR: quic: fix return of quic_dgram_read - MINOR: quic: add config parse source file - MINOR: quic: implement Retry TLS AEAD tag generation - MEDIUM: quic: implement Initial token parsing - MINOR: quic: define retry_source_connection_id TP - MEDIUM: quic: implement Retry emission - MINOR: quic: free xprt tasklet on its thread - BUG/MEDIUM: connection: properly leave stopping list on error - MINOR: pools: enable pools with DEBUG_FAIL_ALLOC as well - MINOR: quic: As server, skip 0-RTT packet number space - MINOR: quic: Do not wakeup the I/O handler before the mux is started - BUG/MEDIUM: htx: Adjust length to add DATA block in an empty HTX buffer - CI: github actions: use cache for OpenTracing - BUG/MINOR: httpclient: don't send an empty body - BUG/MINOR: httpclient: set default Accept and User-Agent headers - BUG/MINOR: httpclient/lua: don't pop the lua stack when getting headers - BUILD/MINOR: fix solaris build with clang. - BUG/MEDIUM: server: avoid changing healthcheck ctx with set server ssl - CI: refactor OpenTracing build script - DOC: management: mark "set server ssl" as deprecated - MEDIUM: cli: yield between each pipelined command - MINOR: channel: add new function co_getdelim() to support multiple delimiters - BUG/MINOR: cli: avoid O(bufsize) parsing cost on pipelined commands - MEDIUM: h2/hpack: emit a Dynamic Table Size Update after settings change - MINOR: quic: Retransmit the TX frames in the same order - MINOR: quic: Remove the packet number space TX MT_LIST - MINOR: quic: Splice the frames which could not be added to packets - MINOR: quic: Add the number of TX bytes to traces - CLEANUP: quic: Replace <nb_pto_dgrams> by <probe> - MINOR: quic: Send two ack-eliciting packets when probing packet number spaces - MINOR: quic: Probe regardless of the congestion control - MINOR: quic: Speeding up handshake completion - MINOR: quic: Release RX Initial packets asap - MINOR: quic: Release asap TX frames to be transmitted - MINOR: quic: Probe even if coalescing - BUG/MEDIUM: cli: Never wait for more data on client shutdown - BUG/MEDIUM: mcli: do not try to parse empty buffers - BUG/MEDIUM: mcli: always realign wrapping buffers before parsing them - BUG/MINOR: stream: make the call_rate only count the no-progress calls - MINOR: quic: do not use quic_conn after dropping it - MINOR: quic: adjust quic_conn refcount decrement - MINOR: quic: fix race-condition on xprt tasklet free - MINOR: quic: free SSL context on quic_conn free - MINOR: quic: Add QUIC_FT_RETIRE_CONNECTION_ID parsing case - MINOR: quic: Wrong packet number space selection - DEBUG: pools: add new build option DEBUG_POOL_INTEGRITY - MINOR: quic: add missing include in quic_sock - MINOR: quic: fix indentation in qc_send_ppkts - MINOR: quic: remove dereferencement of connection when possible - MINOR: quic: set listener accept cb on parsing - MEDIUM: quic/ssl: add new ex data for quic_conn - MINOR: quic: initialize ssl_sock_ctx alongside the quic_conn - MINOR: ssl: fix build in release mode - MINOR: pools: partially uninline pool_free() - MINOR: pools: partially uninline pool_alloc() - MINOR: pools: prepare POOL_EXTRA to be split into multiple extra fields - MINOR: pools: extend pool_cache API to pass a pointer to a caller - DEBUG: pools: add new build option DEBUG_POOL_TRACING - DEBUG: cli: add a new "debug dev fd" expert command - MINOR: fd: register the write side of the poller pipe as well - CI: github actions: use cache for SSL libs - BUILD: debug/cli: condition test of O_ASYNC to its existence - BUILD: pools: fix build error on DEBUG_POOL_TRACING - MINOR: quic: refactor header protection removal - MINOR: quic: handle app data according to mux/connection layer status - MINOR: quic: refactor app-ops initialization - MINOR: receiver: define a flag for local accept - MEDIUM: quic: flag listener for local accept - MINOR: quic: do not manage connection in xprt snd_buf - MINOR: quic: remove wait handshake/L6 flags on init connection - MINOR: listener: add flags field - MINOR: quic: define QUIC flag on listener - MINOR: quic: create accept queue for QUIC connections - MINOR: listener: define per-thr struct - MAJOR: quic: implement accept queue - CLEANUP: mworker: simplify mworker_free_child() - BUILD/DEBUG: lru: update the standalone code to support the revision - DEBUG: lru: use a xorshift generator in the testing code - BUG/MAJOR: compiler: relax alignment constraints on certain structures - BUG/MEDIUM: fd: always align fdtab[] to 64 bytes - MINOR: quic: No DCID length for datagram context - MINOR: quic: Comment fix about the token found in Initial packets - MINOR: quic: Get rid of a struct buffer in quic_lstnr_dgram_read() - MINOR: quic: Remove the QUIC haproxy server packet parser - MINOR: quic: Add new defintion about DCIDs offsets - MINOR: quic: Add a list to QUIC sock I/O handler RX buffer - MINOR: quic: Allocate QUIC datagrams from sock I/O handler - MINOR: proto_quic: Allocate datagram handlers - MINOR: quic: Pass CID as a buffer to quic_get_cid_tid() - MINOR: quic: Convert quic_dgram_read() into a task - CLEANUP: quic: Remove useless definition - MINOR: proto_quic: Wrong allocations for TX rings and RX bufs - MINOR: quic: Do not consume the RX buffer on QUIC sock i/o handler side - MINOR: quic: Do not reset a full RX buffer - MINOR: quic: Attach all the CIDs to the same connection - MINOR: quic: Make usage of by datagram handler trees - MEDIUM: da: new optional data file download scheduler service. - MEDIUM: da: update doc and build for new scheduler mode service. - MEDIUM: da: update module to handle schedule mode. - MINOR: quic: Drop Initial packets with wrong ODCID - MINOR: quic: Wrong RX buffer tail handling when no more contiguous data - MINOR: quic: Iterate over all received datagrams - MINOR: quic: refactor quic CID association with threads - BUG/MEDIUM: resolvers: Really ignore trailing dot in domain names - DEV: flags: Add missing flags - BUG/MINOR: sink: Use the right field in appctx context in release callback - MINOR: sock: move the unused socket cleaning code into its own function - BUG/MEDIUM: mworker: close unused transferred FDs on load failure - BUILD: atomic: make the old HA_ATOMIC_LOAD() support const pointers - BUILD: cpuset: do not use const on the source of CPU_AND/CPU_ASSIGN - BUILD: checks: fix inlining issue on set_srv_agent_[addr,port} - BUILD: vars: avoid overlapping field initialization - BUILD: server-state: avoid using not-so-portable isblank() - BUILD: mux_fcgi: avoid aliasing of a const struct in traces - BUILD: tree-wide: mark a few numeric constants as explicitly long long - BUILD: tools: fix warning about incorrect cast with dladdr1() - BUILD: task: use list_to_mt_list() instead of casting list to mt_list - BUILD: mworker: include tools.h for platforms without unsetenv() - BUG/MINOR: mworker: fix a FD leak of a sockpair upon a failed reload - MINOR: mworker: set the master side of ipc_fd in the worker to -1 - MINOR: mworker: allocate and initialize a mworker_proc - CI: Consistently use actions/checkout@v2 - REGTESTS: Remove REQUIRE_VERSION=1.8 from all tests - MINOR: mworker: sets used or closed worker FDs to -1 - MINOR: quic: Try to accept 0-RTT connections - MINOR: quic: Do not try to treat 0-RTT packets without started mux - MINOR: quic: Do not try to accept a connection more than one time - MINOR: quic: Initialize the connection timer asap - MINOR: quic: Do not use connection struct xprt_ctx too soon - Revert "MINOR: mworker: sets used or closed worker FDs to -1" - BUILD: makefile: avoid testing all -Wno-* options when not needed - BUILD: makefile: validate support for extra warnings by batches - BUILD: makefile: only compute alternative options if required - DEBUG: fd: make sure we never try to insert/delete an impossible FD number - MINOR: mux-quic: add comment - MINOR: mux-quic: properly initialize qcc flags - MINOR: mux-quic: do not consider CONNECTION_CLOSE for the moment - MINOR: mux-quic: create a timeout task - MEDIUM: mux-quic: delay the closing with the timeout - MINOR: mux-quic: release idle conns on process stopping - MINOR: listener: replace the listener's spinlock with an rwlock - BUG/MEDIUM: listener: read-lock the listener during accept() - MINOR: mworker/cli: set expert/experimental mode from the CLI	2022-02-01 18:06:59 +01:00
William Lallemand	7267f78ebe	MINOR: mworker/cli: set expert/experimental mode from the CLI Allow to set the master CLI in expert or experimental mode. No command within the master are unlocked yet, but it gives the ability to send expert or experimental commands to the workers. echo "@1; experimental-mode on; del server be1/s2" \| socat /var/run/haproxy.master - echo "experimental-mode on; @1 del server be1/s2" \| socat /var/run/haproxy.master -	2022-02-01 17:33:06 +01:00
Willy Tarreau	fed93d367c	BUG/MEDIUM: listener: read-lock the listener during accept() Listeners might be disabled by other threads while running in listener_accept() due to a stopping condition or possibly a rebinding error after a failed stop/start. When this happens, the listener's FD is -1 and accesses made by the lower layers to fdtab[-1] do not end up well. This can occasionally be noticed if running at high connection rates in master-worker mode when compiled with ASAN and hammered with 10 reloads per second. From time to time an out-of-bounds error will be reported. One approach could consist in keeping a copy of critical information such as the FD before proceeding but that's not correct since in case of close() the FD might be reassigned to another connection for example. In fact what is needed is to read-lock the listener during this operation so that it cannot change while we're touching it. Tests have shown that using a spinlock only does generally work well but it doesn't scale much with threads and we can see listener_accept() eat 10-15% CPU on a 24 thread machine at 300k conn/s. For this reason the lock was turned to an rwlock by previous commit and this patch only takes the read lock to make sure other operations do not change the listener's state while threads are accepting connections. With this approach, no performance loss was noticed at all and listener_accept() doesn't appear in perf top. This ought to be backported to about all branches that make use of the unlocked listeners, but in practice it seems to mostly concern 2.3 and above, since 2.2 and older will take the FD in the argument (and the race exists there, this FD could end up being reassigned in parallel but there's not much that can be done there to prevent that race; at least a permanent error will be reported). For backports, the current approach is preferred, with a preliminary backport of previous commit "MINOR: listener: replace the listener's spinlock with an rwlock". However if for any reason this commit cannot be backported, the current patch can be modified to simply take a spinlock (tested and works), it will just impact high performance workloads (like DDoS protection).	2022-02-01 16:51:55 +01:00
Willy Tarreau	08b6f96452	MINOR: listener: replace the listener's spinlock with an rwlock We'll need to lock the listener a little bit more during accept() and tests show that a spinlock is a massive performance killer, so let's first switch to an rwlock for this lock. This patch might have to be backported for the next patch to work, and if so, the change is almost mechanical (look for LISTENER_LOCK), but do not forget about the few HA_SPIN_INIT() in the file. There's no reference to this lock outside of listener.c nor listener-t.h.	2022-02-01 16:51:55 +01:00
Amaury Denoyelle	0e0969d6cf	MINOR: mux-quic: release idle conns on process stopping Implement the idle frontend connection cleanup for QUIC mux. Each connection is registered on the mux_stopping_list. On process closing, the mux is notified via a new function qc_wake. This function immediatly release the connection if the parent proxy is stopped. This allows to quickly close the process even if there is QUIC connection stucked on timeout.	2022-02-01 15:42:32 +01:00
Amaury Denoyelle	1136e9243a	MEDIUM: mux-quic: delay the closing with the timeout Do not close immediatly the connection if there is no bidirectional stream opened. Schedule instead the mux timeout when this condition is verified. On the timer expiration, the mux/connection can be freed.	2022-02-01 15:19:35 +01:00
Amaury Denoyelle	aebe26f8ba	MINOR: mux-quic: create a timeout task This task will be used to schedule a timer when there is no activity on the mux. The timeout is set via the "timeout client" from the configuration file. The timeout task process schedule the timeout only on specific conditions. Currently, it's done if there is no opened bidirectional stream. For now this task is not used. This will be implemented in the following commit.	2022-02-01 15:19:35 +01:00
Amaury Denoyelle	d975148776	MINOR: mux-quic: do not consider CONNECTION_CLOSE for the moment Remove the condition on CONNECTION_CLOSE reception to close immediately streams. It can cause some crash as the QUIC xprt layer still access the qcs to send data and handle ACK. The whole interface and buffering between QUIC xprt and mux must be properly reorganized to better handle this case. Once this is done, it may have some sense to free the qcs streams on CONNECTION_CLOSE reception.	2022-02-01 15:19:35 +01:00
Amaury Denoyelle	ce1f30dac8	MINOR: mux-quic: properly initialize qcc flags Set qcc.flags to 0 on qc_init.	2022-02-01 15:19:35 +01:00
Amaury Denoyelle	6a4aebfbfc	MINOR: mux-quic: add comment Explain the qc_release_detached_streams function purpose and interface. Most notably the return code which is the count of released streams.	2022-02-01 10:56:43 +01:00
Willy Tarreau	9aa324de2d	DEBUG: fd: make sure we never try to insert/delete an impossible FD number It's among the cases that would provoke memory corruption, let's add some tests against negative FDs and those larger than the table. This must never ever happen and would currently result in silent corruption or a crash. Better have a noticeable one exhibiting the call chain if that were to happen.	2022-01-31 21:00:35 +01:00
William Lallemand	ce672844dd	Revert "MINOR: mworker: sets used or closed worker FDs to -1" This reverts commit `ea7371e934`. This can't work correctly as we need this FD in the worker to be inserted in the fdtab. The correct way to do it would be to cleanup the mworker_proc in the master after the fork().	2022-01-31 19:06:07 +01:00
Frédéric Lécaille	7fbb94da8d	MINOR: quic: Do not use connection struct xprt_ctx too soon In fact the xprt_ctx of the connection is first stored into quic_conn struct as soon as it is initialized from qc_conn_alloc_ssl_ctx(). As quic_conn_init_timer() is run after this function, we can associate the timer context of the timer to the one from the quic_conn struct.	2022-01-31 16:40:23 +01:00
Frédéric Lécaille	789413caf0	MINOR: quic: Initialize the connection timer asap We must move this initialization from xprt_start() callback, which comes too late (after handshake completion for 1RTT session). This timer must be usable as soon as we have packets to send/receive. Let's initialize it after the TLS context is initialized in qc_conn_alloc_ssl_ctx(). This latter function initializes I/O handler task (quic_conn_io_cb) to send/receive packets.	2022-01-31 16:40:23 +01:00
Frédéric Lécaille	91f083a365	MINOR: quic: Do not try to accept a connection more than one time We add a new flag to mark a connection as already enqueued for acception. This is useful for 0-RTT session where a connection is first enqueued for acception as soon as 0-RTT RX secrets could be derived. Then as for any other connection, we could accept one more time this connection after handshake completion which lead to very bad side effects. Thank you to Amaury for this nice patch.	2022-01-31 16:40:23 +01:00
Frédéric Lécaille	298931d177	MINOR: quic: Do not try to treat 0-RTT packets without started mux We proceed the same was as for 1-RTT packets: we do not try to treat them until the mux is started.	2022-01-31 16:40:23 +01:00
Frédéric Lécaille	61b851d748	MINOR: quic: Try to accept 0-RTT connections When a listener managed to derive 0-RTT RX secrets we consider it accepted the early data. So we enqueue the connection into the accept queue.	2022-01-31 16:40:23 +01:00
William Lallemand	ea7371e934	MINOR: mworker: sets used or closed worker FDs to -1 mworker_cli_sockpair_new() is used to create the socketpair CLI listener of the worker. Its FD is referenced in the mworker_proc structure, however, once it's assigned to the listener the reference should be removed so we don't use it accidentally. The same must be done in case of errors if the FDs were already closed.	2022-01-31 11:10:34 +01:00
William Lallemand	56be0e0146	MINOR: mworker: allocate and initialize a mworker_proc mworker_proc_new() allocates and initializes correctly a mworker_proc structure.	2022-01-28 23:52:36 +01:00
William Lallemand	7e01878e45	MINOR: mworker: set the master side of ipc_fd in the worker to -1 Once the child->ipc_fd[0] is closed in the worker, set the value to -1 so we don't reference a closed FD anymore.	2022-01-28 23:52:26 +01:00
William Lallemand	55a921c914	BUG/MINOR: mworker: fix a FD leak of a sockpair upon a failed reload When starting HAProxy in master-worker, the master pre-allocate a struct mworker_proc and do a socketpair() before the configuration parsing. If the configuration loading failed, the FD are never closed because they aren't part of listener, they are not even in the fdtab. This patch fixes the issue by cleaning the mworker_proc structure that were not asssigned a process, and closing its FDs. Must be backported as far as 2.0, the srv_drop() only frees the memory and could be dropped since it's done before an exec().	2022-01-28 23:47:43 +01:00
Willy Tarreau	4c943fd60b	BUILD: mworker: include tools.h for platforms without unsetenv() In this case we fall back to my_unsetenv() thus we need tools.h to avoid a warning.	2022-01-28 19:04:02 +01:00
Willy Tarreau	cc5cd5b8d8	BUILD: task: use list_to_mt_list() instead of casting list to mt_list There were a few casts of list* to mt_list* that were upsetting some old compilers (not sure about the effect on others). We had created list_to_mt_list() purposely for this, let's use it instead of applying this cast.	2022-01-28 19:04:02 +01:00
Willy Tarreau	f3d5c4b032	BUILD: tools: fix warning about incorrect cast with dladdr1() dladdr1() is used on glibc and takes a void, but we pass it a const ElfW(Sym) and some compilers complain that we're aliasing. Let's just set a may_alias attribute on the local variable to address this. There's no need to backport this unless warnings are reported on older distros or uncommon compilers.	2022-01-28 19:04:02 +01:00
Willy Tarreau	8f0b4e97e7	BUILD: tree-wide: mark a few numeric constants as explicitly long long At a few places in the code the switch/case ond flags are tested against 64-bit constants without explicitly being marked as long long. Some 32-bit compilers complain that the constant is too large for a long, and other likely always use long long there. Better fix that as it's uncertain what others which do not complain do. It may be backported to avoid doubts on uncommon platforms if needed, as it touches very few areas.	2022-01-28 19:04:02 +01:00
Willy Tarreau	31a8306b93	BUILD: mux_fcgi: avoid aliasing of a const struct in traces fcgi_trace() declares fconn as a const and casts its mbuf array to (struct buffer*), which rightfully upsets some older compilers. Better just declare it as a writable variable and get rid of the cast. It's harmless anyway. This has been there since 2.1 with commit `5c0f859c2` ("MINOR: mux-fcgi/trace: Register a new trace source with its events") and doens't need to be backported though it would not harm either.	2022-01-28 19:04:02 +01:00
Willy Tarreau	74bc991600	BUILD: server-state: avoid using not-so-portable isblank() Once in a while we get rid of this one. isblank() is missing on old C libraries and only matches two values, so let's just replace it. It was brought with this commit in 2.4: `0bf268e18` ("MINOR: server: Be more strict on the server-state line parsing") It may be backported though it's really not important.	2022-01-28 19:04:02 +01:00
Willy Tarreau	e90dde1edf	BUILD: vars: avoid overlapping field initialization Compiling vars.c with gcc 4.2 shows that we're initializing some local structs field members in a not really portable way: src/vars.c: In function 'vars_parse_cli_set_var': src/vars.c:1195: warning: initialized field overwritten src/vars.c:1195: warning: (near initialization for 'px.conf.args') src/vars.c:1195: warning: initialized field overwritten src/vars.c:1195: warning: (near initialization for 'px.conf') src/vars.c:1201: warning: initialized field overwritten src/vars.c:1201: warning: (near initialization for 'rule.conf') It's totally harmless anyway, but better clean this up.	2022-01-28 19:04:02 +01:00
Willy Tarreau	95d3eaff36	BUILD: checks: fix inlining issue on set_srv_agent_[addr,port} These functions are declared as external functions in check.h and as inline functions in check.c. Let's move them as static inline in check.h. This appeared in 2.4 with the following commits: `4858fb2e1` ("MEDIUM: check: align agentaddr and agentport behaviour") `1c921cd74` ("BUG/MINOR: check: consitent way to set agentaddr") While harmless (it only triggers build warnings with some gcc 4.x), it should probably be backported where the paches above are present to keep the code consistent.	2022-01-28 19:04:02 +01:00
Willy Tarreau	a65b4933ba	BUILD: cpuset: do not use const on the source of CPU_AND/CPU_ASSIGN The man page indicates that CPU_AND() and CPU_ASSIGN() take a variable, not a const on the source, even though it doesn't make much sense. But with older libcs, this triggers a build warning: src/cpuset.c: In function 'ha_cpuset_and': src/cpuset.c:53: warning: initialization discards qualifiers from pointer target type src/cpuset.c: In function 'ha_cpuset_assign': src/cpuset.c:101: warning: initialization discards qualifiers from pointer target type Better stick stricter to the documented API as this is really harmless here. There's no need to backport it (unless build issues are reported, which is quite unlikely).	2022-01-28 19:04:02 +01:00
Willy Tarreau	e08acaed19	BUG/MEDIUM: mworker: close unused transferred FDs on load failure When the master process is reloaded on a new config, it will try to connect to the previous process' socket to retrieve all known listening FDs to be reused by the new listeners. If listeners were removed, their unused FDs are simply closed. However there's a catch. In case a socket fails to bind, the master will cancel its startup and swithc to wait mode for a new operation to happen. In this case it didn't close the possibly remaining FDs that were left unused. It is very hard to hit this case, but it can happen during a troubleshooting session with fat fingers. For example, let's say a config runs like this: frontend ftp bind 1.2.3.4:20000-29999 The admin wants to extend the port range down to 10000-29999 and by mistake ends up with: frontend ftp bind 1.2.3.41:20000-29999 Upon restart the bind will fail if the address is not present, and the master will then switch to wait mode without releasing the previous FDs for 1.2.3.4:20000-29999 since they're now apparently unused. Then once the admin fixes the config and does: frontend ftp bind 1.2.3.4:10000-29999 The service will start, but will bind new sockets, half of them overlapping with the previous ones that were not properly closed. This may result in a startup error (if SO_REUSEPORT is not enabled or not available), in a FD number exhaustion (if the error is repeated many times), or in connections being randomly accepted by the process if they sometimes land on the old FD that nobody listens on. This patch will need to be backported as far as 1.8, and depends on previous patch: MINOR: sock: move the unused socket cleaning code into its own function Note that before 2.3 most of the code was located inside haproxy.c, so the patch above should probably relocate the function there instead of sock.c.	2022-01-28 19:04:02 +01:00
Willy Tarreau	b510116fd2	MINOR: sock: move the unused socket cleaning code into its own function The startup code used to scan the list of unused sockets retrieved from an older process, and to close them one by one. This also required that the knowledge of the internal storage of these temporary sockets was known from outside sock.c and that the code was copy-pasted at every call place. This patch moves this into sock.c under the name sock_drop_unused_old_sockets(), and removes the xfer_sock_list definition from sock.h since the rest of the code doesn't need to know this. This cleanup is minimal and preliminary to a future fix that will need to be backported to all versions featuring FD transfers over the CLI.	2022-01-28 19:04:02 +01:00
Christopher Faulet	dd0b144c3a	BUG/MINOR: sink: Use the right field in appctx context in release callback In the release callback, ctx.peers was used instead of ctx.sft. Concretly, it is not an issue because the appctx context is an union and these both fields are structures with a unique pointer. But it will be a problem if that changes. This patch must be backported as far as 2.2.	2022-01-28 17:56:18 +01:00
Christopher Faulet	0a82cf4c16	BUG/MEDIUM: resolvers: Really ignore trailing dot in domain names When a string is converted to a domain name label, the trailing dot must be ignored. In resolv_str_to_dn_label(), there is a test to do so. However, the trailing dot is not really ignored. The character itself is not copied but the string index is still moved to the next char. Thus, this trailing dot is counted in the length of the last encoded part of the domain name. Worst, because the copy is skipped, a garbage character is included in the domain name. This patch should fix the issue #1528. It must be backported as far as 2.0.	2022-01-28 17:56:18 +01:00
Amaury Denoyelle	0442efd214	MINOR: quic: refactor quic CID association with threads Do not use an extra DCID parameter on new_quic_cid to be able to associated a new generated CID to a thread ID. Simply do the computation inside the function. The API is cleaner this way. This also has the effects to improve the apparent randomness of CIDs. With the previous version the first byte of all CIDs are identical for a connection which could lead to privacy issue. This version may not be totally perfect on this aspect but it improves the situation.	2022-01-28 16:29:27 +01:00
Frédéric Lécaille	df1c7c78c1	MINOR: quic: Iterate over all received datagrams Make the listener datagram handler iterate over all received datagrams	2022-01-28 16:08:07 +01:00
Frédéric Lécaille	1712b1df59	MINOR: quic: Wrong RX buffer tail handling when no more contiguous data The producer must know where is the tailing hole in the RX buffer when it purges it from consumed datagram. This is done allocating a fake datagram with the remaining number of bytes which cannot be produced at the tail of the RX buffer as length.	2022-01-28 16:08:07 +01:00
Frédéric Lécaille	dc36404c36	MINOR: quic: Drop Initial packets with wrong ODCID According to the RFC 9000, the client ODCID must have a minimal length of 8 bytes.	2022-01-28 16:08:07 +01:00
Frédéric Lécaille	74904a4792	MINOR: quic: Make usage of by datagram handler trees The CID trees are no more attached to the listener receiver but to the underlying datagram handlers (one by thread) which run always on the same thread. So, any operation on these trees do not require any locking.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	9ea9463d47	MINOR: quic: Attach all the CIDs to the same connection We copy the first octet of the original destination connection ID to any CID for the connection calling new_quic_cid(). So this patch modifies only this function to take a dcid as passed parameter.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	320744b53d	MINOR: quic: Do not reset a full RX buffer As the RX buffer is not consumed by the sock i/o handler as soon as a datagram is produced, when full an RX buffer must not be reset. The remaining room is consumed without modifying it. The consumer has a represention of its contents: a list of datagrams.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	37ae505c21	MINOR: quic: Do not consume the RX buffer on QUIC sock i/o handler side Rename quic_lstnr_dgram_read() to quic_lstnr_dgram_dispatch() to reflect its new role. After calling this latter, the sock i/o handler must consume the buffer only if the datagram it received is detected as wrong by quic_lstnr_dgram_dispatch(). The datagram handler task mark the datagram as consumed atomically setting ->buf to NULL value. The sock i/o handler is responsible of flushing its RX buffer before using it. It also keeps a datagram among the consumed ones so that to pass it to quic_lstnr_dgram_dispatch() and prevent it from allocating a new one.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	794d068d8f	MINOR: proto_quic: Wrong allocations for TX rings and RX bufs As mentionned in the comment, the tx_qrings and rxbufs members of receiver struct must be pointers to pointers! Modify the functions responsible of their allocations consequently. Note that this code could work because sizeof rxbuf and sizeof tx_qrings are greater than the size of pointer!	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	25bc8875d7	MINOR: quic: Convert quic_dgram_read() into a task quic_dgram_read() parses all the QUIC packets from a UDP datagram. It is the best candidate to be converted into a task, because is processing data unit is the UDP datagram received by the QUIC sock i/o handler. If correct, this datagram is added to the context of a task, quic_lstnr_dghdlr(), a conversion of quic_dgram_read() into a task. This task pop a datagram from an mt_list and passes it among to the packet handler (quic_lstnr_pkt_rcv()). Modify the quic_dgram struct to play the role of the old quic_dgram_ctx struct when passed to quic_lstnr_pkt_rcv(). Modify the datagram handlers allocation to set their tasks to quic_lstnr_dghdlr().	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	220894a5d6	MINOR: quic: Pass CID as a buffer to quic_get_cid_tid() Very minor modification so that this function might be used for a context without CID (at datagram level).	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	69dd5e6a0b	MINOR: proto_quic: Allocate datagram handlers Add quic_dghdlr new struct do define datagram handler tasks, one by thread. Allocate them and attach them to the listener receiver part calling quic_alloc_dghdlrs_listener() newly implemented function.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	3d4bfe708a	MINOR: quic: Allocate QUIC datagrams from sock I/O handler Add quic_dgram new structure to store information about datagrams received by the sock I/O handler (quic_sock_fd_iocb) and its associated pool. Implement quic_get_dgram_dcid() to retrieve the datagram DCID which must be the same for all the packets in the datagram. Modify quic_lstnr_dgram_read() called by the sock I/O handler to allocate a quic_dgram each time a correct datagram is found and add it to the sock I/O handler rxbuf dgram list.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	53898bba81	MINOR: quic: Add a list to QUIC sock I/O handler RX buffer This list will be used to store datagrams in the rxbuf struct used by the quic_sock_fd_iocb() QUIC sock I/O handler with one rxbuf by thread.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	9cc64e2dba	MINOR: quic: Remove the QUIC haproxy server packet parser This function is no more used anymore, broken and uses code shared with the listener packet parser. This is becoming anoying to continue to modify it without testing each time we modify the code it shares with the listener packet parser.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	3d55462654	MINOR: quic: Get rid of a struct buffer in quic_lstnr_dgram_read() This is to be sure xprt functions do not manipulate the buffer struct passed as parameter to quic_lstnr_dgram_read() from low level datagram I/O callback in quic_sock.c (quic_sock_fd_iocb()).	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	055ee6c14b	MINOR: quic: Comment fix about the token found in Initial packets Mention that the token is sent only by servers in both server and listener packet parsers. Remove a "TO DO" section in listener packet parser because there is nothing more to do in this function about the token	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	4852101fd2	MINOR: quic: No DCID length for datagram context This quic_dgram_ctx struct member is used to denote if we are parsing a new datagram (null value), or a coalesced packet into the current datagram (non null value). But it was never set.	2022-01-27 16:37:55 +01:00
Willy Tarreau	97ea9c49f1	BUG/MEDIUM: fd: always align fdtab[] to 64 bytes There's a risk that fdtab is not 64-byte aligned. The first effect is that it may cause false sharing between cache lines resulting in contention when adjacent FDs are used by different threads. The second is related to what is explained in commit "BUG/MAJOR: compiler: relax alignment constraints on certain structures", i.e. that modern compilers might make use of aligned vector operations to zero some entries, and would crash. We do not use any memset() or so on fdtab, so the risk is almost inexistent, but that's not a reason for violating some valid assumptions. This patch addresses this by allocating 64 extra bytes and aligning the structure manually (this is an extremely cheap solution for this specific case). The original address is stored in a new variable "fdtab_addr" and is the one that gets freed. This remains extremely simple and should be easily backportable. A dedicated aligned allocator later would help, of course. This needs to be backported as far as 2.2. No issue related to this was reported yet, but it could very well happen as compilers evolve. In addition this should preserve high performance across restarts (i.e. no more dependency on allocator's alignment).	2022-01-27 16:28:10 +01:00
Willy Tarreau	8e92738ffd	DEBUG: lru: use a xorshift generator in the testing code The standalone testing code used to rely on rand(), but switching to a xorshift generator speeds up the test by 7% which is important to accurately measure the real impact of the LRU code itself.	2022-01-27 16:28:10 +01:00
Willy Tarreau	bf9c07fd91	BUILD/DEBUG: lru: update the standalone code to support the revision The standalone testing code didn't implement the revision and didn't build anymore, let's fix that.	2022-01-27 16:28:10 +01:00
William Lallemand	08cb945a9b	CLEANUP: mworker: simplify mworker_free_child() Remove useless checks and simplify the function.	2022-01-27 15:33:40 +01:00
Amaury Denoyelle	cfa2d5648f	MAJOR: quic: implement accept queue Do not proceed to direct accept when creating a new quic_conn. Wait for the QUIC handshake to succeeds to insert the quic_conn in the accept queue. A tasklet is then woken up to call listener_accept to accept the quic_conn. The most important effect is that the connection/mux layers are not instantiated at the same time as the quic_conn. This forces to delay some process to be sure that the mux is allocated : * initialization of mux transport parameters * installation of the app-ops Also, the mux instance is not checked now to wake up the quic_conn tasklet. This is safe because the xprt-quic code is now ready to handle the absence of the connection/mux layers. Note that this commit has a deep impact as it changes significantly the lower QUIC architecture. Most notably, it breaks the 0-RTT feature.	2022-01-26 16:13:54 +01:00
Amaury Denoyelle	f68b2cb816	MINOR: listener: define per-thr struct Create a new structure li_per_thread. This is uses as an array in the listener structure, with an entry allocated per thread. The new function li_init_per_thr is responsible of the allocation. For now, li_per_thread contains fields only useful for QUIC listeners. As such, it is only allocated for QUIC listeners.	2022-01-26 16:13:54 +01:00
Amaury Denoyelle	2ce99fe4bf	MINOR: quic: create accept queue for QUIC connections Create a new type quic_accept_queue to handle QUIC connections accept. A queue will be allocated for each thread. It contains a list of listeners which contains at least one quic_conn ready to be accepted and the tasklet to run listener_accept for these listeners.	2022-01-26 16:13:51 +01:00
Amaury Denoyelle	b59b88950a	MINOR: quic: define QUIC flag on listener Mark QUIC listeners with the flag LI_F_QUIC_LISTENER. It is set by the proto-quic layer on the add listener callback. This allows to override more clearly the accept callback on quic_session_accept.	2022-01-26 15:25:45 +01:00
Amaury Denoyelle	cbe090d42f	MINOR: quic: remove wait handshake/L6 flags on init connection The connection is allocated after finishing the QUIC handshake. Remove handshake/L6 flags when initializing the connection as handshake is finished with success at this stage.	2022-01-26 15:25:45 +01:00
Amaury Denoyelle	9fa15e5413	MINOR: quic: do not manage connection in xprt snd_buf Remove usage of connection in quic_conn_from_buf. As connection and quic_conn are decorrelated, it is not logical to check connection flags when using sendto. This require to store the L4 peer address in quic_conn to be able to use sendto. This change is required to delay allocation of connection.	2022-01-26 15:25:38 +01:00
Amaury Denoyelle	683b5fc7b8	MEDIUM: quic: flag listener for local accept QUIC connections are distributed accross threads by xprt-quic according to their CIDs. As such disable the thread selection in listener_accept for QUIC listeners. This prevents connection from migrating to another threads after its allocation which can results in unexpected side-effects.	2022-01-26 11:59:12 +01:00
Amaury Denoyelle	7f7713d6ef	MINOR: receiver: define a flag for local accept This flag is named RX_F_LOCAL_ACCEPT. It will be activated for special receivers where connection balancing to threads is already handle outside of listener_accept, such as with QUIC listeners.	2022-01-26 11:22:20 +01:00
Amaury Denoyelle	4b40f19f92	MINOR: quic: refactor app-ops initialization Add a new function in mux-quic to install app-ops. For now this functions is called during the ALPN negotiation of the QUIC handshake. This change will be useful when the connection accept queue will be implemented. It will be thus required to delay the app-ops initialization because the mux won't be allocated anymore during the QUIC handshake.	2022-01-26 10:59:33 +01:00
Amaury Denoyelle	0b1f93127f	MINOR: quic: handle app data according to mux/connection layer status Define a new enum to represent the status of the mux/connection layer above a quic_conn. This is important to know if it's possible to handle application data, or if it should be buffered or dropped.	2022-01-26 10:57:17 +01:00
Amaury Denoyelle	8ae28077b9	MINOR: quic: refactor header protection removal Adjust the function to check if header protection can be removed. It can now be used both for a single packet in qc_lstnr_pkt_rcv and in the quic_conn handler to handle buffered packets for a specific encryption level.	2022-01-26 10:51:16 +01:00
Willy Tarreau	f70fdde591	BUILD: pools: fix build error on DEBUG_POOL_TRACING When squashing commit `add43fa43` ("DEBUG: pools: add new build option DEBUG_POOL_TRACING") I managed to break the build and to fail to detect it even after the rebase and a full rebuild :-(	2022-01-25 15:59:18 +01:00
Willy Tarreau	410942b92a	BUILD: debug/cli: condition test of O_ASYNC to its existence David Carlier reported a build breakage on Haiku since commit `5be7c198e` ("DEBUG: cli: add a new "debug dev fd" expert command") due to O_ASYNC not being defined. Ilya also reported it broke the build on Cygwin. It's not that portable and sometimes defined as O_NONBLOCK for portability. But here we don't even need that, as we already condition other flags, let's just ignore it if it does not exist.	2022-01-25 14:51:53 +01:00
Willy Tarreau	3a6af1e5e8	MINOR: fd: register the write side of the poller pipe as well The poller's pipe was only registered on the read side since we don't need to poll to write on it. But this leaves some known FDs so it's better to also register the write side with no event. This will allow to show them in "show fd" and to avoid dumping them as unhandled FDs. Note that the only other type of unhandled FDs left are: - stdin/stdout/stderr - epoll FDs The later can be registered upon startup though but at least a dummy handler would be needed to keep the fdtab clean.	2022-01-24 20:41:25 +01:00
Willy Tarreau	5be7c198e5	DEBUG: cli: add a new "debug dev fd" expert command This command will scan the whole file descriptors space to look for existing FDs that are unknown to haproxy's fdtab, and will try to dump a maximum number of information about them (including type, mode, device, size, uid/gid, cloexec, O_* flags, socket types and addresses when relevant). The goal is to help detecting inherited FDs from parent processes as well as potential leaks. Some of those listed are actually known but handled so deep into some systems that they're not in the fdtab (such as epoll FDs or inter- thread pipes). This might be refined in the future so that these ones become known and do not appear. Example of output: $ socat - /tmp/sock1 <<< "expert-mode on;debug dev fd" 0 type=tty. mod=0620 dev=0x8803 siz=0 uid=1000 gid=5 fs=0x16 ino=0x6 getfd=+0 getfl=O_RDONLY,O_APPEND 1 type=tty. mod=0620 dev=0x8803 siz=0 uid=1000 gid=5 fs=0x16 ino=0x6 getfd=+0 getfl=O_RDONLY,O_APPEND 2 type=tty. mod=0620 dev=0x8803 siz=0 uid=1000 gid=5 fs=0x16 ino=0x6 getfd=+0 getfl=O_RDONLY,O_APPEND 3 type=pipe mod=0600 dev=0 siz=0 uid=1000 gid=100 fs=0xc ino=0x18112348 getfd=+0 4 type=epol mod=0600 dev=0 siz=0 uid=0 gid=0 fs=0xd ino=0x3674 getfd=+0 getfl=O_RDONLY 33 type=pipe mod=0600 dev=0 siz=0 uid=1000 gid=100 fs=0xc ino=0x24af8251 getfd=+0 getfl=O_RDONLY 34 type=epol mod=0600 dev=0 siz=0 uid=0 gid=0 fs=0xd ino=0x3674 getfd=+0 getfl=O_RDONLY 36 type=pipe mod=0600 dev=0 siz=0 uid=1000 gid=100 fs=0xc ino=0x24af8d1b getfd=+0 getfl=O_RDONLY 37 type=epol mod=0600 dev=0 siz=0 uid=0 gid=0 fs=0xd ino=0x3674 getfd=+0 getfl=O_RDONLY 39 type=pipe mod=0600 dev=0 siz=0 uid=1000 gid=100 fs=0xc ino=0x24afa04f getfd=+0 getfl=O_RDONLY 41 type=pipe mod=0600 dev=0 siz=0 uid=1000 gid=100 fs=0xc ino=0x24af8252 getfd=+0 getfl=O_RDONLY 42 type=epol mod=0600 dev=0 siz=0 uid=0 gid=0 fs=0xd ino=0x3674 getfd=+0 getfl=O_RDONLY	2022-01-24 20:26:09 +01:00
Willy Tarreau	add43fa43e	DEBUG: pools: add new build option DEBUG_POOL_TRACING This new option, when set, will cause the callers of pool_alloc() and pool_free() to be recorded into an extra area in the pool that is expected to be helpful for later inspection (e.g. in core dumps). For example it may help figure that an object was released to a pool with some sub-fields not yet released or that a use-after-free happened after releasing it, with an immediate indication about the exact line of code that released it (possibly an error path). This only works with the per-thread cache, and even objects refilled from the shared pool directly into the thread-local cache will have a NULL there. That's not an issue since these objects have not yet been freed. It's worth noting that pool_alloc_nocache() continues not to set any caller pointer (e.g. when the cache is empty) because that would require a possibly undesirable API change. The extra cost is minimal (one pointer per object) and this completes well with DEBUG_POOL_INTEGRITY.	2022-01-24 16:40:48 +01:00
Willy Tarreau	0e2a5b4b61	MINOR: pools: extend pool_cache API to pass a pointer to a caller This adds a caller to pool_put_to_cache() and pool_get_from_cache() which will optionally be used to pass a pointer to their callers. For now it's not used, only the API is extended to support this pointer.	2022-01-24 16:40:48 +01:00
Willy Tarreau	d392973dcc	MINOR: pools: partially uninline pool_alloc() The pool_alloc() function was already a wrapper to __pool_alloc() which was also inlined but took a set of flags. This latter was uninlined and moved to pool.c, and pool_alloc()/pool_zalloc() turned to macros so that they can more easily evolve to support debugging options. The number of call places made this code grow over time and doing only this change saved ~1% of the whole executable's size.	2022-01-24 16:40:48 +01:00
Willy Tarreau	15c322c413	MINOR: pools: partially uninline pool_free() The pool_free() function has become a bit big over time due to the extra consistency checks. It used to remain inline only to deal cleanly with the NULL pointer free that's quite present on some structures (e.g. in stream_free()). Here we're splitting the function in two: - __pool_free() does the inner block without the pointer test and becomes a function ; - pool_free() is now a macro that only checks the pointer and calls __pool_free() if needed. The use of a macro versus an inline function is only motivated by an easier intrumentation of the code later. With this change, the code size reduces by ~1%, which means that at this point all pool_free() call places used to represent more than 1% of the total code size.	2022-01-24 16:40:48 +01:00
Amaury Denoyelle	7c564bfdd3	MINOR: ssl: fix build in release mode Fix potential null pointer dereference. In fact, this case is not possible, only a mistake in SSL ex-data initialization may cause it : either connection is set or quic_conn, which allows to retrieve the bind_conf. A BUG_ON was already present but this does not cover release build.	2022-01-24 11:15:48 +01:00
Amaury Denoyelle	33ac346ba8	MINOR: quic: initialize ssl_sock_ctx alongside the quic_conn Extract the allocation of ssl_sock_ctx from qc_conn_init to a dedicated function qc_conn_alloc_ssl_ctx. This function is called just after allocating a new quic_conn, without waiting for the initialization of the connection. It allocates the ssl_sock_ctx and the quic_conn tasklet. This change is now possible because the SSL callbacks are dealing with a quic_conn instance. This change is required to be able to delay the connection allocation and handle handshake packets without it.	2022-01-24 10:30:49 +01:00
Amaury Denoyelle	9320dd5385	MEDIUM: quic/ssl: add new ex data for quic_conn Allow to register quic_conn as ex-data in SSL callbacks. A new index is used to identify it as ssl_qc_app_data_index. Replace connection by quic_conn as SSL ex-data when initializing the QUIC SSL session. When using SSL callbacks in QUIC context, the connection is now NULL. Used quic_conn instead to retrieve the required parameters. Also clean up The same changes are conducted inside the QUIC SSL methods of xprt-quic : connection instance usage is replaced by quic_conn.	2022-01-24 10:30:49 +01:00
Amaury Denoyelle	57af069571	MINOR: quic: set listener accept cb on parsing Define a special accept cb for QUIC listeners to quic_session_accept(). This operation is conducted during the proto.add callback when creating listeners. A special care is now taken care when setting the standard callback session_accept_fd() to not overwrite if already defined by the proto layer.	2022-01-24 10:30:49 +01:00
Amaury Denoyelle	29632b8b10	MINOR: quic: remove dereferencement of connection when possible Some functions of xprt-quic were still using connection instead of quic_conn. This must be removed as the two are decorrelated : a quic_conn can exist without a connection.	2022-01-24 10:30:49 +01:00
Amaury Denoyelle	74f2292557	MINOR: quic: fix indentation in qc_send_ppkts Adjust wrong mixing of tabs/spaces.	2022-01-24 10:30:49 +01:00
Amaury Denoyelle	4d29504c58	MINOR: quic: add missing include in quic_sock Add quic_sock.h include in corresponding source file quic_sock.c.	2022-01-24 10:30:49 +01:00
Willy Tarreau	0575d8fd76	DEBUG: pools: add new build option DEBUG_POOL_INTEGRITY When enabled, objects picked from the cache are checked for corruption by comparing their contents against a pattern that was placed when they were inserted into the cache. Objects are also allocated in the reverse order, from the oldest one to the most recent, so as to maximize the ability to detect such a corruption. The goal is to detect writes after free (or possibly hardware memory corruptions). Contrary to DEBUG_UAF this cannot detect reads after free, but may possibly detect later corruptions and will not consume extra memory. The CPU usage will increase a bit due to the cost of filling/checking the area and for the preference for cold cache instead of hot cache, though not as much as with DEBUG_UAF. This option is meant to be usable in production.	2022-01-21 19:07:48 +01:00
Frédéric Lécaille	39ba1c3e12	MINOR: quic: Wrong packet number space selection It is possible that the listener is in INITIAL state, but have to probe with Handshake packets. In this case, when entering qc_prep_pkts() there is nothing to do. We must select the next packet number space (or encryption level) to be able to probe with such packet type.	2022-01-21 17:38:11 +01:00
Frédéric Lécaille	2cca241780	MINOR: quic: Add QUIC_FT_RETIRE_CONNECTION_ID parsing case At this time, we do not do anything. This is only to prevent a packet from being parsed and to pass some test irrespective of the CIDs management.	2022-01-21 17:38:11 +01:00
Amaury Denoyelle	2d9794b03a	MINOR: quic: free SSL context on quic_conn free Free the SSL context attached to the quic_conn when freeing the connection. This fixes a memory leak for every QUIC connection.	2022-01-21 15:20:07 +01:00
Amaury Denoyelle	760da3be57	MINOR: quic: fix race-condition on xprt tasklet free Remove the unsafe call to tasklet_free in quic_close. At this stage the tasklet may already be scheduled by an other threads even after if the quic_conn refcount is now null. It will probably cause a crash on the next tasklet processing. Use tasklet_kill instead to ensure that the tasklet is freed in a thread-safe way. Note that quic_conn_io_cb is not protected by the refcount so only the quic_conn pinned thread must kill the tasklet.	2022-01-21 15:19:31 +01:00
Amaury Denoyelle	2eb7b30715	MINOR: quic: adjust quic_conn refcount decrement Adjust slightly refcount code decrement on quic_conn close. A new function named quic_conn_release is implemented. This function is responsible to remove the quic_conn from CIDs trees and decrement the refcount to free the quic_conn once all threads have finished to work with it. For now, quic_close is responsible to call it so the quic_conn is scheduled to be free by upper layers. In the future, it may be useful to delay it to be able to send remaining data or waiting for missing ACKs for example. This simplify quic_conn_drop which do not require the lock anymore. Also, this can help to free the connection more quickly in some cases.	2022-01-21 15:03:17 +01:00
Amaury Denoyelle	9c4da93796	MINOR: quic: do not use quic_conn after dropping it quic_conn_drop decrement the refcount and may free the quic_conn if reaching 0. The quic_conn should not be dereferenced again after it in any case even for traces.	2022-01-21 15:02:56 +01:00
Willy Tarreau	6c539c4b8c	BUG/MINOR: stream: make the call_rate only count the no-progress calls We have an anti-looping protection in process_stream() that detects bugs that used to affect a few filters like compression in the past which sometimes forgot to handle a read0 or a particular error, leaving a thread looping at 100% CPU forever. When such a condition is detected, an alert it emitted and the process is killed so that it can be replaced by a sane one: [ALERT] (19061) : A bogus STREAM [0x274abe0] is spinning at 2057156 calls per second and refuses to die, aborting now! Please report this error to developers [strm=0x274abe0,3 src=unix fe=MASTER be=MASTER dst=<MCLI> txn=(nil),0 txn.req=-,0 txn.rsp=-,0 rqf=c02000 rqa=10000 rpf=88000021 rpa=8000000 sif=EST,40008 sib=DIS,84018 af=(nil),0 csf=0x274ab90,8600 ab=0x272fd40,1 csb=(nil),0 cof=0x25d5d80,1300:PASS(0x274aaf0)/RAW((nil))/unix_stream(9) cob=(nil),0:NONE((nil))/NONE((nil))/NONE(0) filters={}] call trace(11): \| 0x4dbaab [c7 04 25 01 00 00 00 00]: stream_dump_and_crash+0x17b/0x1b4 \| 0x4df31f [e9 bd c8 ff ff 49 83 7c]: process_stream+0x382f/0x53a3 (...) One problem with this detection is that it used to only count the call rate because we weren't sure how to make it more accurate, but the threshold was high enough to prevent accidental false positives. There is actually one case that manages to trigger it, which is when sending huge amounts of requests pipelined on the master CLI. Some short requests such as "show version" are sufficient to be handled extremely fast and to cause a wake up of an analyser to parse the next request, then an applet to handle it, back and forth. But this condition is not an error, since some data are being forwarded by the stream, and it's easy to detect it. This patch modifies the detection so that update_freq_ctr() only applies to calls made without CF_READ_PARTIAL nor CF_WRITE_PARTIAL set on any of the channels, which really indicates that nothing is happening at all. This is greatly sufficient and extremely effective, as the call above is still caught (shutr being ignored by an analyser) while a loop on the master CLI now has no effect. The "call_rate" field in the detailed "show sess" output will now be much lower, except for bogus streams, which may help spot them. This field is only there for developers anyway so it's pretty fine to slightly adjust its meaning. This patch could be backported to stable versions in case of reports of such an issue, but as that's unlikely, it's not really needed.	2022-01-20 18:56:57 +01:00
Willy Tarreau	a4e4d66f70	BUG/MEDIUM: mcli: always realign wrapping buffers before parsing them Pipelined commands easily result in request buffers to wrap, and the master-cli parser only deals with linear buffers since it needs contiguous keywords to look for in a list. As soon as a buffer wraps, some commands are ignored and the parser is called in loops because the wrapped data do not leave the buffer. Let's take the easiest path that's already used at the HTTP layer, we simply realign the buffer if its input wraps. This rarely happens anyway (typically once per buffer), remains reasonably cheap and guarantees this cannot happen anymore. This needs to be backported as far as 2.0.	2022-01-20 18:56:57 +01:00
Willy Tarreau	6cd93f52e9	BUG/MEDIUM: mcli: do not try to parse empty buffers When pcli_parse_request() is called with an empty buffer, it still tries to parse it and can go on believing it finds an empty request if the last char before the beginning of the buffer is a '\n'. In this case it overwrites it with a zero and processes it as an empty command, doing nothing but not making the buffer progress. This results in an infinite loop that is stopped by the watchdog. For a reason related to another issue (yet to be fixed), this can easily be reproduced by pipelining lots of commands such as "show version". Let's add a length check after the search for a '\n'. This needs to be backported as far as 2.0.	2022-01-20 18:56:57 +01:00
Christopher Faulet	0f727dabf5	BUG/MEDIUM: cli: Never wait for more data on client shutdown When a shutdown is detected on the cli, we try to execute all pending commands first before closing the connection. It is required because commands execution is serialized. However, when the last part is a partial command, the cli connection is not closed, waiting for more data. Because there is no timeout for now on the cli socket, the connection remains infinitely in this state. And because the maxconn is set to 10, if it happens several times, the cli socket quickly becomes unresponsive because all its slots are waiting for more data on a closed connections. This patch should fix the issue #1512. It must be backported as far as 2.0.	2022-01-20 18:56:39 +01:00
Frédéric Lécaille	94fca87f6a	MINOR: quic: Probe even if coalescing Again, we fix a reminiscence of the way we probed before probing by packet. When we were probing by datagram we inspected <prv_pkt> to know if we were coalescing several packets. There is no need to do that at all when probing by packet. Furthermore this could lead to blocking situations where we want to probe but are limited by the congestion control (<cwnd> path variable). This must not be the case. When probing we must do it regardless of the congestion control.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	e87524d41c	MINOR: quic: Release asap TX frames to be transmitted This is done only for ack-eliciting frames to be sent from Initial and Handshake packet number space when discarding them.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	a6255f53e8	MINOR: quic: Release RX Initial packets asap This is to free up some space in the RX buffer as soon as possible.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	04e63aa6ef	MINOR: quic: Speeding up handshake completion If a client resend Initial CRYPTO data, this is because it did not receive all the server Initial CRYPTO data. With this patch we prepare a fast retransmission without waiting for the PTO timer expiration sending old Initial CRYPTO data, coalescing them with Handshake CRYPTO if present in the same datagram. Furthermore we send also a datagram made of previously sent Hanshashke CRYPTO data if any.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	f4e5a7c644	MINOR: quic: Probe regardless of the congestion control When probing, we must not take into an account the congestion control window. This was not completely correctly implemented: qc_build_frms() could fail because of this limit when comparing the head of the packet againts the congestion control window. With this patch we make it fail only when we are not probing.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	0fa553d0c2	MINOR: quic: Send two ack-eliciting packets when probing packet number spaces This is to avoid too much PTO timer expirations for 01RTT and Handshake packet number spaces. Furthermore we are not limited by the anti-amplication for 01RTT packet number space. According to the RFC we can send up to two packets.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	ce6602d887	CLEANUP: quic: Replace <nb_pto_dgrams> by <probe> This modification should have come with this commit: "MINOR: quic: Remove nb_pto_dgrams quic_conn struct member" where the nb_pto_dgrams quic_conn struct member was removed.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	8b6ea17105	MINOR: quic: Add the number of TX bytes to traces This should be helpful to diagnose some issues regarding packet loss and recovery issues.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	cba4cd427e	MINOR: quic: Splice the frames which could not be added to packets When building packets to send, we build frames computing their sizes to have more chance to be added to new packets. There are rare cases where this packet coult not be built because of the congestion control which may for instance prevent us from building a packet with padding (retransmitted Initial packets). In such a case, the pre-built frames were lost because added to the packet frame list but not move packet to the packet number space they come from. With this patch we add the frames to the packet only if it could be built and move them back to the packet number space if not.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	82468ea98e	MINOR: quic: Remove the packet number space TX MT_LIST There is no need to use an MT_LIST to store frames to send from a packet number space. This is a reminiscence for multi-threading support for the TX part.	2022-01-20 16:43:06 +01:00
Frédéric Lécaille	7065dd0895	MINOR: quic: Retransmit the TX frames in the same order This is only to please the peer. We resend the TX frames in the same order they have been sent.	2022-01-20 16:35:43 +01:00
Willy Tarreau	39a0a1e120	MEDIUM: h2/hpack: emit a Dynamic Table Size Update after settings change As reported by @jinsubsim in github issue #1498, there is an interoperability issue between nghttp2 as a client and a few servers among which haproxy (in fact likely all those which do not make use of the dynamic headers table in responses or which do not intend to use a larger table), when reducing the header table size below 4096. These are easily testable this way: nghttp -v -H":method: HEAD" --header-table-size=0 https://$SITE It will result in a compression error for those which do not start with an HPACK dynamic table size update opcode. There is a possible interpretation of the H2 and HPACK specs that says that an HPACK encoder must send an HPACK headers table update confirming the new size it will be using after having acknowledged it, because since it's possible for a decoder to advertise a late SETTINGS and change it after transfers have begun, the initially advertised value might very well be seen as a first change from the initial setting, and the HPACK spec doesn't specify the side which causes the change that triggers a DTSU update, which was essentially summed up in this question from nghttp2's author when this issue was already raised 6 years ago, but which didn't really find a solid response by then: https://lists.w3.org/Archives/Public/ietf-http-wg/2015OctDec/0107.html The ongoing consensus based on what some servers are doing and that aims at limiting interoperability issues seems to be that a DTSU is expected for each reduction from the current size, which should be reflected in the next revision of the H2 spec: https://github.com/httpwg/http2-spec/pull/1005 Given that we do not make use of this table we can emit a DTSU of zero before encoding any HPACK frame. However, some clients do not support receiving DTSU with such values (e.g. VTest) so we cannot do it inconditionnally! The current patch aims at sticking as close to the spec as possible by proceeding this way: - when a SETTINGS_HEADER_TABLE_SIZE is received, a flag is set indicating that the value changed - before sending any HPACK frame, this flag is checked to see if an update is wanted and if none was sent - in this case a DTSU of size zero is emitted and a flag is set to mention it was emitted so that it never has to be sent again This addresses the problem with nghttp2 without affecting VTest. More context is available here: https://github.com/nghttp2/nghttp2/issues/1660 https://lists.w3.org/Archives/Public/ietf-http-wg/2021OctDec/0235.html Many thanks to @jinsubsim for this report and participating to the issue that led to an improvement of the H2 spec. This should be backported to stable releases in a timely manner, ideally as far as 2.4 once the h2spec update is merged, then to other versions after a few months of observation or in case an issue around this is reported.	2022-01-20 05:01:03 +01:00
Willy Tarreau	0011c25144	BUG/MINOR: cli: avoid O(bufsize) parsing cost on pipelined commands Sending pipelined commands on the CLI using a semi-colon as a delimiter has a cost that grows linearly with the buffer size, because co_getline() is called for each word and looks up a '\n' in the whole buffer while copying its contents into a temporary buffer. This causes huge parsing delays, for example 3s for 100k "show version" versus 110ms if parsed only once for a default 16k buffer. This patch makes use of the new co_getdelim() function to support both an LF and a semi-colon as delimiters so that it's no more needed to parse the whole buffer, and that commands are instantly retrieved. We still need to rely on co_getline() in payload mode as escapes and semi-colons are not used there. It should likely be backported where CLI processing speed matters, but will require to also backport previous patch "MINOR: channel: add new function co_getdelim() to support multiple delimiters". It's worth noting that backporting it without "MEDIUM: cli: yield between each pipelined command" would significantly increase the ratio of disconnections caused by empty request buffers, for the sole reason that the currently slow parsing grants more time to request data to come in. As such it would be better to backport the patch above before taking this one.	2022-01-19 19:16:47 +01:00
Willy Tarreau	c514365317	MINOR: channel: add new function co_getdelim() to support multiple delimiters For now we have co_getline() which reads a buffer and stops on LF, and co_getword() which reads a buffer and stops on one arbitrary delimiter. But sometimes we'd need to stop on a set of delimiters (CR and LF, etc). This patch adds a new function co_getdelim() which takes a set of delimiters as a string, and constructs a small map (32 bytes) that's looked up during parsing to stop after the first delimiter found within the set. It also supports an optional escape character that skips a delimiter (typically a backslash). For the rest it works exactly like the two other variants.	2022-01-19 19:16:47 +01:00
Willy Tarreau	fa7b4f6691	MEDIUM: cli: yield between each pipelined command Pipelining commands on the CLI is sometimes needed for batched operations such as map deletion etc, but it causes two problems: - some possibly long-running commands will be run in series without yielding, possibly causing extremely long latencies that will affect quality of service and even trigger the watchdog, as seen in github issue #1515. - short commands that end on a buffer size boundary, when not run in interactive mode, will often cause the socket to be closed when the last command is parsed, because the buffer is empty. This patch proposes a small change to this: by yielding in the CLI applet after processing a command when there are data left, we significantly reduce the latency, since only one command is executed per call, and we leave an opportunity for the I/O layers to refill the request buffer with more commands, hence to execute all of them much more often. With this change there's no more watchdog triggered on long series of "del map" on large map files, and the operations are much less disturbed. It would be desirable to backport this patch to stable versions after some period of observation in recent versions.	2022-01-19 19:16:47 +01:00
William Dauchy	a087f87875	BUG/MEDIUM: server: avoid changing healthcheck ctx with set server ssl While giving a fresh try to `set server ssl` (which I wrote), I realised the behavior is a bit inconsistent. Indeed when using this command over a server with ssl enabled for the data path but also for the health check path we have: - data and health check done using tls - emit `set server be_foo/srv0 ssl off` - data path and health check path becomes plain text - emit `set server be_foo/srv0 ssl on` - data path becomes tls and health check path remains plain text while I thought the end result would be: - data path and health check path comes back in tls In the current code we indeed erase all connections while deactivating, but restore only the data path while activating. I made this mistake in the past because I was testing with a case where the health check plain text by default. There are several ways to solve this issue. The cleanest one would probably be to avoid changing the health check connection when we use `set server ssl` command, and create a new command `set server ssl-check` to change this. For now I assumed this would be ok to simply avoid changing the health check path and be more consistent. This patch tries to address that and also update the documentation. It should not break the existing usage with health check on plain text, as in this case they should have `no-check-ssl` in defaults. Without this patch, it makes the command unusable in an env where you have a list of server to add along the way with initial `server-template`, and all using tls for data and healthcheck path. For 2.6 we should probably reconsider and add `set server ssl-check` command for better granularity of cases. If this solution is accepted, this patch should be backported up to >= 2.4. The alternative solution was to restore the previous state, but I believe this will create even more confusion in the future. Signed-off-by: William Dauchy <wdauchy@gmail.com>	2022-01-18 12:05:17 +01:00
William Lallemand	01e2be84d7	BUG/MINOR: httpclient/lua: don't pop the lua stack when getting headers hlua_httpclient_table_to_hdrs() does a lua_pop(L, 1) at the end of the function, this is supposed to be done in the caller and it is already be done in hlua_httpclient_send(). This call has the consequence of poping the next parameter of the httpclient, ignoring it. This patch fixes the issue by removing the lua_pop(L, 1). Must be backported in 2.5.	2022-01-14 20:51:31 +01:00
William Lallemand	bad9c8cac4	BUG/MINOR: httpclient: set default Accept and User-Agent headers Some servers require at least an Accept and a User-Agent header in the request. This patch sets some default value. Must be backported in 2.5.	2022-01-14 20:46:21 +01:00
William Lallemand	e1e045f4d7	BUG/MINOR: httpclient: don't send an empty body Forbid the httpclient to send an empty chunked client when there is no data to send. It does happen when doing a simple GET too. Must be backported in 2.5.	2022-01-14 18:10:43 +01:00
Christopher Faulet	28e7ba8688	BUG/MEDIUM: htx: Adjust length to add DATA block in an empty HTX buffer htx_add_data() is able to partially consume data. However there is a bug when the HTX buffer is empty. The data length is not properly adjusted. Thus, if it exceeds the HTX buffer size, no block is added. To fix the issue, the length is now adjusted first. This patch must be backported as far as 2.0.	2022-01-13 09:34:22 +01:00
Frédéric Lécaille	b80b20c6ff	MINOR: quic: Do not wakeup the I/O handler before the mux is started If we wakeup the I/O handler before the mux is started, it is possible it has enough time to parse the ClientHello TLS message and update the mux transport parameters, leading to a crash. So, we initialize ->qcc quic_conn struct member at the very last time, when the mux if fully initialized. The condition to wakeup the I/O handler from lstnr_rcv_pkt() is: xprt context and mux both initialized. Note that if the xprt context is initialized, it implies its tasklet is initialized. So, we do not check anymore this latter condition.	2022-01-12 18:08:29 +01:00
Frédéric Lécaille	bec186dde5	MINOR: quic: As server, skip 0-RTT packet number space This is true only when we are building packets. A QUIC server never sends 0-RTT packets. So let't skip the associated TLS encryption level.	2022-01-12 18:08:29 +01:00
Willy Tarreau	3b990fe0be	BUG/MEDIUM: connection: properly leave stopping list on error The stopping-list management introduced by commit `d3a88c1c3` ("MEDIUM: connection: close front idling connection on soft-stop") missed two error paths in the H1 and H2 muxes. The effect is that if a stream or HPACK table couldn't be allocated for these incoming connections, we would leave with the connection freed still attached to the stopping_list and it would never leave it, resulting in use-after-free hence either a crash or a data corruption. This is marked as medium as it only happens under extreme memory pressure or when playing with tune.fail-alloc. Other stability issues remain in such a case so that abnormal behaviors cannot be explained by this bug alone. This must be backported to 2.4.	2022-01-12 17:31:01 +01:00
Amaury Denoyelle	9ab2fb3921	MINOR: quic: free xprt tasklet on its thread Free the ssl_sock_ctx tasklet in quic_close() instead of quic_conn_drop(). This ensures that the tasklet is destroyed safely by the same thread. This has no impact as the free operation was previously conducted with care and should not be responsible of any crash.	2022-01-12 15:21:27 +01:00
Amaury Denoyelle	b76ae69513	MEDIUM: quic: implement Retry emission Implement the emission of Retry packets. These packets are emitted in response to Initial from clients without token. The token from the Retry packet contains the ODCID from the Initial packet. By default, Retry packet emission is disabled and the handshake can continue without address validation. To enable Retry, a new bind option has been defined named "quic-force-retry". If set, the handshake must be conducted only after receiving a token in the Initial packet.	2022-01-12 11:08:48 +01:00
Amaury Denoyelle	c3b6f4d484	MINOR: quic: define retry_source_connection_id TP Define a new QUIC transport parameter retry_source_connection_id. This parameter is set only by server, after issuing a Retry packet.	2022-01-12 11:08:48 +01:00
Amaury Denoyelle	5ff1c9778c	MEDIUM: quic: implement Initial token parsing Implement the parsing of token from Initial packets. It is expected that the token contains a CID which is the DCID from the Initial packet received from the client without token which triggers a Retry packet. This CID is then used for transport parameters. Note that at the moment Retry packet emission is not implemented. This will be achieved in a following commit.	2022-01-12 11:08:48 +01:00
Amaury Denoyelle	6efec292ef	MINOR: quic: implement Retry TLS AEAD tag generation Implement a new QUIC TLS related function quic_tls_generate_retry_integrity_tag(). This function can be used to calculate the AEAD tag of a Retry packet.	2022-01-12 11:08:48 +01:00
Amaury Denoyelle	c8b4ce4a47	MINOR: quic: add config parse source file Create a new dedicated source file for QUIC related options parsing on the bind line.	2022-01-12 11:08:48 +01:00
Amaury Denoyelle	ce340fe4a7	MINOR: quic: fix return of quic_dgram_read It is expected that quic_dgram_read() returns the total number of bytes read. Fix the return value when the read has been successful. This bug has no impact as in the end the return value is not checked by the caller.	2022-01-12 11:08:48 +01:00
Frédéric Lécaille	1aa57d32bb	MINOR: quic: Do not dereference ->conn quic_conn struct member ->conn quic_conn struct member is a connection struct object which may be released from several places. With this patch we do our best to stop dereferencing this member as much as we can.	2022-01-12 09:49:49 +01:00
Frédéric Lécaille	ba85acdc70	MINOR: quid: Add traces quic_close() and quic_conn_io_cb() This is to have an idea of possible remaining issues regarding the connection terminations.	2022-01-11 16:56:04 +01:00
Frédéric Lécaille	81cd3c8eed	MINOR: quic: Wrong CRYPTO frame concatenation This commit was not correct: "MINOR: quic: Only one CRYPTO frame by encryption level" Indeed, when receiving CRYPTO data from TLS stack for a packet number space, there are rare cases where there is already other frames than CRYPTO data frames in the packet number space, especially for 01RTT packet number space. This is very often with quant as client.	2022-01-11 16:12:31 +01:00
Frédéric Lécaille	2fe8b3be20	MINOR: quic: Flag the connection as being attached to a listener We do not rely on connection objects to know if we are a listener or not.	2022-01-11 16:12:31 +01:00
Frédéric Lécaille	19cd46e6e5	MINOR: quic: Reset ->conn quic_conn struct member when calling qc_release() There may be remaining locations where ->conn quic_conn struct member is used. So let's reset this. Add a trace to have an idead when this connection is released.	2022-01-11 16:12:31 +01:00
Frédéric Lécaille	5f7f118b31	MINOR: quic: Remaining TRACEs with connection as firt arg This is a quic_conn struct which is expected by TRACE_()* macros	2022-01-11 16:12:31 +01:00
David CARLIER	bb10dad5a8	BUILD: cpuset: fix build issue on macos introduced by previous change The build on macos was broken by recent commit `df91cbd58` ("MINOR: cpuset: switch to sched_setaffinity for FreeBSD 14 and above."), let's move the variable declaration inside the ifdef.	2022-01-11 15:09:49 +01:00
Christopher Faulet	b4eca0e908	BUG/MAJOR: mux-h1: Don't decrement .curr_len for unsent data A regression was introduced by commit `140f1a58` ("BUG/MEDIUM: mux-h1: Fix splicing by properly detecting end of message"). To detect end of the outgoing message, when the content-length is announced, we count amount of data already sent. But only data really sent must be counted. If the output buffer is full, we can fail to send data (fully or partially). In this case, we must take care to only count sent data. Otherwise we may think too much data were sent and an internal error may be erroneously reported. This patch should fix issues #1510 and #1511. It must be backported as far as 2.4.	2022-01-11 09:15:13 +01:00
Remi Tricot-Le Breton	a996763619	BUG/MINOR: ssl: Store client SNI in SSL context in case of ClientHello error If an error is raised during the ClientHello callback on the server side (ssl_sock_switchctx_cbk), the servername callback won't be called and the client's SNI will not be saved in the SSL context. But since we use the SSL_get_servername function to return this SNI in the ssl_fc_sni sample fetch, that means that in case of error, such as an SNI mismatch with a frontend having the strict-sni option enabled, the sample fetch would not work (making strict-sni related errors hard to debug). This patch fixes that by storing the SNI as an ex_data in the SSL context in case the ClientHello callback returns an error. This way the sample fetch can fallback to getting the SNI this way. It will still first call the SSL_get_servername function first since it is the proper way of getting a client's SNI when the handshake succeeded. In order to avoid memory allocations are runtime into this highly used runtime function, a new memory pool was created to store those client SNIs. Its entry size is set to 256 bytes since SNIs can't be longer than 255 characters. This fixes GitHub #1484. It can be backported in 2.5.	2022-01-10 16:31:22 +01:00
William Lallemand	f82afbb9cd	BUG/MEDIUM: mworker: don't use _getsocks in wait mode Since version 2.5 the master is automatically re-executed in wait-mode when the config is successfully loaded, puting corner cases of the wait mode in plain sight. When using the -x argument and with the right timing, the master will try to get the FDs again in wait mode even through it's not needed anymore, which will harm the worker by removing its listeners. However, if it fails, (and it's suppose to, sometimes), the master will exit with EXIT_FAILURE because it does not have the MODE_MWORKER flag, but only the MODE_MWORKER_WAIT flag. With the consequence of killing the workers. This patch fixes the issue by restricting the use of _getsocks to some modes. This patch must be backported in every version supported, even through the impact should me more harmless in version prior to 2.5.	2022-01-07 18:44:27 +01:00
Frédéric Lécaille	99942d6f4c	MINOR: quic: Non-optimal use of a TX buffer When full, after having reset the writer index, let's reuse the TX buffer in any case.	2022-01-07 17:58:26 +01:00
Frédéric Lécaille	f010f0aaf2	MINOR: quic: Missing retransmission from qc_prep_fast_retrans() In fact we must look for the first packet with some ack-elicting frame to in the packet number space tree to retransmit from. Obviously there may be already retransmit packets which are not deemed as lost and still present in the packet number space tree for TX packets.	2022-01-07 17:58:26 +01:00
Frédéric Lécaille	d4ecf94827	MINOR: quic: Only one CRYPTO frame by encryption level When receiving CRYPTO data from the TLS stack, concatenate the CRYPTO data to the first allocated CRYPTO frame if present. This reduces by one the number of handshake packets built for a connection with a standard size certificate.	2022-01-07 17:58:26 +01:00
Ilya Shipitsin	37d3e38130	CLEANUP: assorted typo fixes in the code and comments This is 30th iteration of typo fixes	2022-01-07 14:42:54 +01:00
David CARLIER	df91cbd584	MINOR: cpuset: switch to sched_setaffinity for FreeBSD 14 and above. Following up previous update on cpuset-t.h. Ultimately, at some point the cpuset_setaffinity code path could be removed.	2022-01-07 06:53:51 +01:00
William Dauchy	a9dd901143	MINOR: proxy: add option idle-close-on-response Avoid closing idle connections if a soft stop is in progress. By default, idle connections will be closed during a soft stop. In some environments, a client talking to the proxy may have prepared some idle connections in order to send requests later. If there is no proper retry on write errors, this can result in errors while haproxy is reloading. Even though a proper implementation should retry on connection/write errors, this option was introduced to support back compat with haproxy < v2.4. Indeed before v2.4, we were waiting for a last request to be able to add a "connection: close" header and advice the client to close the connection. In a real life example, this behavior was seen in AWS using the ALB in front of a haproxy. The end result was ALB sending 502 during haproxy reloads. This patch was tested on haproxy v2.4, with a regular reload on the process, and a constant trend of requests coming in. Before the patch, we see regular 502 returned to the client; when activating the option, the 502 disappear. This patch should help fixing github issue #1506. In order to unblock some v2.3 to v2.4 migraton, this patch should be backported up to v2.4 branch. Signed-off-by: William Dauchy <wdauchy@gmail.com> [wt: minor edits to the doc to mention other options to care about] Signed-off-by: Willy Tarreau <w@1wt.eu>	2022-01-06 09:09:51 +01:00
Frédéric Lécaille	6b6631593f	MINOR: quic: Re-arm the PTO timer upon datagram receipt When block by the anti-amplification limit, this is the responsability of the client to unblock it sending new datagrams. On the server side, even if not well parsed, such datagrams must trigger the PTO timer arming.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	078634d126	MINOR: quic: PTO timer too often reset It must be reset when the anti-amplication was reached but only if the peer address was not validated.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	41a076087b	MINOR: quic: Flag asap the connection having reached the anti-amplification limit The best location to flag the connection is just after having built the packet which reached the anti-amplication limit.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	de6f7c503e	MINOR: quic: Prepare Handshake packets asap after completed handshake Switch back to QUIC_HS_ST_SERVER_HANDSHAKE state after a completed handshake if acks must be send. Also ensure we build post handshake frames only one time without using prev_st variable and ensure we discard the Handshake packet number space only one time.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	917a7dbdc7	MINOR: quic: Do not drop secret key but drop the CRYPTO data We need to be able to decrypt late Handshake packets after the TLS secret keys have been discarded. If not the peer send Handshake packet which have not been acknowledged. But for such packets, we discard the CRYPTO data.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	ee2b8b377f	MINOR: quic: Improve qc_prep_pkts() flexibility We want to be able to choose the encryption levels to be used by qc_prep_pkts() outside of it.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	f7ef97698a	MINOR: quic: Comment fix. When we drop a packet with unknown length, this is the entire datagram which must be skipped.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	a56054e438	MINOR: quic: Probe several packet number space upon timer expiration When the loss detection timer expires, we SHOULD include new data in our probing packets (RFC 9002 par 6.2.4. Sending Probe Packets).	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	db6a4727cf	MINOR: quic: Probe Initial packet number space more often Especially when the PTO expires for Handshake packet number space and when Initial packets are still flying (for QUIC servers).	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	3bb457c4ba	MINOR: quic: Speeding up Handshake Completion According to RFC 9002 par. 6.2.3. when receving duplicate Initial CRYPTO data a server may a packet containing non unacknowledged before the PTO expiry.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	63556772cc	MINOR: quic: qc_prep_pkts() code moving Move the switch default case code out of the switch to improve the readibily.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	5732062cd2	MINOR: quic: Useless test in qc_prep_pkts() These tests were there to initiate PTO probing but they are not correct. Furthermore they may break the PTO probing process and lead to useless packet building.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	dd51da599e	MINOR: quic: Wrong packet number space trace in qc_prep_pkts() It was always the first packet number space information which was dumped.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	466e9da145	MINOR: quic: Remove nb_pto_dgrams quic_conn struct member For now on we rely on tx->pto_probe pktns struct member to inform the packet building function we want to probe.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	22576a2e55	MINOR: quic: Wrong ack_delay compution before calling quic_loss_srtt_update() RFC 9002 5.3. Estimating smoothed_rtt and rttvar: MUST use the lesser of the acknowledgment delay and the peer's max_ack_delay after the handshake is confirmed.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	dc90c07715	MINOR: quic: Wrong loss time computation in qc_packet_loss_lookup() This part as been modified by the RFC since our first implementation.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	22cfd83890	MINOR: quic: Add trace about in flight bytes by packet number space This parameter is useful to diagnose packet loss detection issues.	2022-01-04 17:30:00 +01:00
Frédéric Lécaille	fde2a98dd1	MINOR: quic: Wrong traces after rework TRACE_*() macros must take a quic_conn struct as first argument.	2022-01-04 17:30:00 +01:00
Christopher Faulet	7bf46bb9a9	BUG/MEDIUM: http-ana: Preserve response's FLT_END analyser on L7 retry When a filter is attached on a stream, the FLT_END analyser must not be removed from the response channel on L7 retry. It is especially important because CF_FLT_ANALYZE flag is still set. This means the synchronization between the two sides when the filter ends can be blocked. Depending on the timing, this can freeze the stream infinitely or lead to a spinning loop. Note that the synchronization between the two sides at the end of the analysis was introduced because the stream was reused in HTTP between two transactions. But, since the HTX was introduced, a new stream is created for each transaction. So it is probably possible to remove this step for 2.2 and higher. This patch must be backported as far as 2.0.	2022-01-04 10:56:04 +01:00
David Carlier	ae5c42f4d0	BUILD/MINOR: tools: solaris build fix on dladdr. dladdr takes a mutable address on this platform.	2022-01-03 14:43:51 +01:00
Ilya Shipitsin	5e87bcf870	CLEANUP: assorted typo fixes in the code and comments This is 29th iteration of typo fixes	2022-01-03 14:40:58 +01:00
Willy Tarreau	1513c5479a	MEDIUM: pools: release cached objects in batches With this patch pool_evict_last_items builds clusters of up to CONFIG_HAP_POOL_CLUSTER_SIZE entries so that accesses to the shared pools are reduced by CONFIG_HAP_POOL_CLUSTER_SIZE and the inter- thread contention is reduced by as much..	2022-01-02 19:35:26 +01:00
Willy Tarreau	43937e920f	MEDIUM: pools: start to batch eviction from local caches Since previous patch we can forcefully evict multiple objects from the local cache, even when evicting basd on the LRU entries. Let's define a compile-time configurable setting to batch releasing of objects. For now we set this value to 8 items per round. This is marked medium because eviction from the LRU will slightly change in order to group the last items that are freed within a single cache instead of accurately scanning only the oldest ones exactly in their order of appearance. But this is required in order to evolve towards batched removals.	2022-01-02 19:35:26 +01:00
Willy Tarreau	a0b5831eed	MEDIUM: pools: centralize cache eviction in a common function We currently have two functions to evict cold objects from local caches: pool_evict_from_local_cache() to evict from a single cache, and pool_evict_from_local_caches() to evict oldest objects from all caches. The new function pool_evict_last_items() focuses on scanning oldest objects from a pool and releasing a predefined number of them, either to the shared pool or to the system. For now they're evicted one at a time, but the next step will consist in creating clusters.	2022-01-02 19:35:26 +01:00
Willy Tarreau	337410c5a4	MINOR: pools: pass the objects count to pool_put_to_shared_cache() This is in order to let the caller build the cluster of items to be released. For now single items are released hence the count is always 1.	2022-01-02 19:35:26 +01:00
Willy Tarreau	148160b027	MINOR: pools: prepare pool_item to support chained clusters In order to support batched allocations and releases, we'll need to prepare chains of items linked together and that can be atomically attached and detached at once. For this we implement a "down" pointer in each pool_item that points to the other items belonging to the same group. For now it's always NULL though freeing functions already check them when trying to release everything.	2022-01-02 19:35:26 +01:00
Willy Tarreau	361e31e3fe	MEDIUM: pool: compute the number of evictable entries once per pool In pool_evict_from_local_cache() we used to check for room left in the pool for each and every object. Now we compute the value before entering the loop and keep into a local list what has to be released, and call the OS-specific functions for the other ones. It should already save some cycles since it's not needed anymore to recheck for the pool's filling status. But the main expected benefit comes from the ability to pre-construct a list of all releasable objects, that will later help with grouping them.	2022-01-02 19:35:26 +01:00
Willy Tarreau	c16ed3b090	MINOR: pool: introduce pool_item to represent shared pool items In order to support batch allocation from/to shared pools, we'll have to support a specific representation for pool objects. The new pool_item structure will be used for this. For now it only contains a "next" pointer that matches exactly the current storage model. The few functions that deal with the shared pool entries were adapted to use the new type. There is no functionality difference at this point.	2022-01-02 19:35:26 +01:00
Willy Tarreau	b46674a283	MINOR: pool: check for pool's fullness outside of pool_put_to_shared_cache() Instead of letting pool_put_to_shared_cache() pass the object to the underlying OS layer when there's no more room, let's have the caller check if the pool is full and either call pool_put_to_shared_cache() or call pool_free_nocache(). Doing this sensibly simplifies the code as this function now only has to deal with a pool and an item and only for cases where there are local caches and shared caches. As the code was simplified and the calls more isolated, the function was moved to pool.c. Note that it's only called from pool_evict_from_local_cache{,s}() and that a part of its logic might very well move there when dealing with batches.	2022-01-02 19:35:26 +01:00
Willy Tarreau	afe2c4a1fc	MINOR: pool: allocate from the shared cache through the local caches One of the thread scaling challenges nowadays for the pools is the contention on the shared caches. There's never any situation where we have a shared cache and no local cache anymore, so we can technically afford to transfer objects from the shared cache to the local cache before returning them to the user via the regular path. This adds a little bit more work per object per miss, but will permit batch processing later. This patch simply moves pool_get_from_shared_cache() to pool.c under the new name pool_refill_local_from_shared(), and this function does not return anything but it places the allocated object at the head of the local cache.	2022-01-02 19:27:57 +01:00
Willy Tarreau	8c4927098e	CLEANUP: pools: get rid of the POOL_LINK macro The POOL_LINK macro is now only used for debugging, and it still requires ifdefs around, which needlessly complicates the code. Let's replace it and the calling code with a new pair of macros: POOL_DEBUG_SET_MARK() and POOL_DEBUG_CHECK_MARK(), that respectively store and check the pool pointer in the extra location at the end of the pool. This removes 4 pairs of ifdefs in the middle of the code.	2022-01-02 12:44:19 +01:00
Willy Tarreau	799f6143ca	CLEANUP: pools: do not use the extra pointer to link shared elements This practice relying on POOL_LINK() dates from the era where there were no pool caches, but given that the structures are a bit more complex now and that pool caches do not make use of this feature, it is totally useless since released elements have already been overwritten, and yet it complicates the architecture and prevents from making simplifications and optimizations. Let's just get rid of this feature. The pointer to the origin pool is preserved though, as it helps detect incorrect frees and serves as a canary for overflows.	2022-01-02 12:44:19 +01:00
Willy Tarreau	d5ec100661	MINOR: pools: always evict oldest objects first in pool_evict_from_local_cache() For an unknown reason, despite the comment stating that we were evicting oldest objects first from the local caches, due to the use of LIST_NEXT, the newest were evicted, since pool_put_to_cache() uses LIST_INSERT(). Some tests on 16 threads show that evicting oldest objects instead can improve performance by 0.5-1% especially when using shared pools.	2022-01-02 12:40:14 +01:00
William Lallemand	e69563fd8e	BUG/MEDIUM: ssl: free the ckch instance linked to a server This patch unlinks and frees the ckch instance linked to a server during the free of this server. This could have locked certificates in a "Used" state when removing servers dynamically from the CLI. And could provoke a segfault once we try to dynamically update the certificate after that. This must be backported as far as 2.4.	2021-12-30 16:56:52 +01:00
William Lallemand	231610ad9c	BUG/MINOR: ssl: free the fields in srv->ssl_ctx A lot of free are missing in ssl_sock_free_srv_ctx(), this could result in memory leaking when removing dynamically a server via the CLI. This must be backported in every branches, by removing the fields that does not exist in the previous branches.	2021-12-30 13:43:04 +01:00
William Lallemand	2c776f1c30	BUG/MEDIUM: ssl: initialize correctly ssl w/ default-server This bug was introduced by `d817dc73` ("MEDIUM: ssl: Load client certificates in a ckch for backend servers") in which the creation of the SSL_CTX for a server was moved to the configuration parser when using a "crt" keyword instead of being done in ssl_sock_prepare_srv_ctx(). The patch `0498fa40` ("BUG/MINOR: ssl: Default-server configuration ignored by server") made it worse by setting the same SSL_CTX for every servers using a default-server. Resulting in any SSL option on a server applied to every server in its backend. This patch fixes the issue by reintroducing a string which store the path of certificate inside the server structure, and loading the certificate in ssl_sock_prepare_srv_ctx() again. This is a quick fix to backport, a cleaner way can be achieve by always creating the SSL_CTX in ssl_sock_prepare_srv_ctx() and splitting properly the ssl_sock_load_srv_cert() function. This patch fixes issue #1488. Must be backported as far as 2.4.	2021-12-29 14:42:16 +01:00
Willy Tarreau	654726db5a	MINOR: debug: add support for -dL to dump library names at boot This is a second help to dump loaded library names late at boot, once external code has already been initialized. The purpose is to provide a format that makes it easy to pass to "tar" to produce an archive containing the executable and the list of dependencies. For example if haproxy is started as "haproxy -f foo.cfg", a config check only will suffice to quit before starting, "-q" will be used to disable undesired output messages, and -dL will be use to dump libraries. This will result in such a command to trivially produce a tarball of loaded libraries: ./haproxy -q -c -dL -f foo.cfg \| tar -T - -hzcf archive.tgz	2021-12-28 17:07:13 +01:00
Willy Tarreau	6ab7b21a11	MINOR: debug: add ability to dump loaded shared libraries Many times core dumps reported by users who experience trouble are difficult to exploit due to missing system libraries. Sometimes, having just a list of loaded libraries and their respective addresses can already provide some hints about some problems. This patch makes a step in that direction by adding a new "show libs" command that will try to enumerate the list of object files that are loaded in memory, relying on the dynamic linker for this. It may also be used to detect that some foreign code embarks other undesired libs (e.g. some external Lua modules). At the moment it's only supported on glibc when USE_DL is set, but it's implemented in a way that ought to make it reasonably easy to be extended to other platforms.	2021-12-28 16:59:00 +01:00
Willy Tarreau	b4ff6f4ae9	BUG/MEDIUM: peers: properly skip conn_cur from incoming messages The approach used for skipping conn_cur in commit `db2ab8218` ("MEDIUM: stick-table: never learn the "conn_cur" value from peers") was wrong, it only works with simple tables but as soon as frequency counters or arrays are exchanged after conn_cur, the stream is desynchronized and incorrect values are read. This is because the fields have a variable length depending on their types and cannot simply be skipped by a "continue" statement. Let's change the approach to make sure we continue to completely parse these local-only fields, and only drop the value at the moment we're about to store them, since this is exactly the intent. A simpler approach could consist in having two sets of stktable_data_ptr() functions, one for retrieval and one for storage, and to make the store function return a NULL pointer for local types. For now this doesn't seem worth the trouble. This fixes github issue #1497. Thanks to @brenc for the reproducer. This must be backported to 2.5.	2021-12-24 13:48:39 +01:00
Willy Tarreau	266d540549	BUG/MEDIUM: backend: fix possible sockaddr leak on redispatch A subtle change of target address allocation was introduced with commit `68cf3959b` ("MINOR: backend: rewrite alloc of stream target address") in 2.4. Prior to this patch, a target address was allocated by function assign_server_address() only if none was previously allocated. After the change, the allocation became unconditional. Most of the time it makes no difference, except when we pass multiple times through connect_server() with SF_ADDR_SET cleared. The most obvious fix would be to avoid allocating that address there when already set, but the root cause is that since introduction of dynamically allocated addresses, the SF_ADDR_SET flag lies. It can be cleared during redispatch or during a queue redistribution without the address being released. This patch instead gives back all its correct meaning to SF_ADDR_SET and guarantees that when not set no address is allocated, by freeing that address at the few places the flag is cleared. The flag could even be removed so that only the address is checked but that would require to touch many areas for no benefit. The easiest way to test it is to send requests to a proxy with l7 retries enabled, which forwards to a server returning 500: defaults mode http timeout client 1s timeout server 1s timeout connect 1s retry-on all-retryable-errors retries 1 option redispatch listen proxy bind *:5000 server app 0.0.0.0:5001 frontend dummy-app bind :5001 http-request return status 500 Issuing "show pools" on the CLI will show that pool "sockaddr" grows as requests are redispatched, and remains stable with the fix. Even "ps" will show that the process' RSS grows by ~160B per request. This fix will need to be backported to 2.4. Note that before 2.5, there's no strm->si[1].dst, strm->target_addr must be used instead. This addresses github issue #1499. Special thanks to Daniil Leontiev for providing a well-documented reproducer.	2021-12-24 11:50:01 +01:00
Amaury Denoyelle	9979d0d1ea	BUG/MINOR: quic: fix potential use of uninit pointer Properly initialized the ssl_sock_ctx pointer in qc_conn_init. This is required to avoid to set an undefined pointer in qc.xprt_ctx if argument *xprt_ctx is NULL.	2021-12-23 16:33:47 +01:00
Amaury Denoyelle	c6fab98f9b	BUG/MINOR: quic: fix potential null dereference This is not a real issue because found_in_dcid can not be set if qc is NULL.	2021-12-23 16:32:19 +01:00
Amaury Denoyelle	76f47caacc	MEDIUM: quic: implement refcount for quic_conn Implement a refcount on quic_conn instance. By default, the refcount is 0. Two functions are implemented to manipulate it. * qc_conn_take() which increments the refcount * qc_conn_drop() which decrements it. If the refcount is 0 BEFORE the substraction, the instance is freed. The refcount is incremented on retrieve_qc_conn_from_cid() or when allocating a new quic_conn in qc_lstnr_pkt_rcv(). It is substracted most notably by the xprt.close operation and at the end of qc_lstnr_pkt_rcv(). The increments/decrements should be conducted under the CID lock to guarantee thread-safety.	2021-12-23 16:06:07 +01:00
Amaury Denoyelle	0a29e13835	MINOR: quic: delete timer task on quic_close() The timer task is attached to the connection-pinned thread. Only this thread can delete it. With the future refcount implementation of quic_conn, every thread can be responsible to remove the quic_conn via quic_conn_free(). Thus, the timer task deletion is moved from the calling function quic_close().	2021-12-23 16:06:07 +01:00
Amaury Denoyelle	e81fed9a54	MINOR: quic: replace usage of ssl_sock_ctx by quic_conn Big refactoring on xprt-quic. A lot of functions were using the ssl_sock_ctx as argument to only access the related quic_conn. All these arguments are replaced by a quic_conn parameter. As a convention, the quic_conn instance is always the first parameter of these functions. This commit is part of the rearchitecture of xprt-quic layers and the separation between xprt and connection instances.	2021-12-23 16:06:06 +01:00
Amaury Denoyelle	741eacca47	MINOR: quic: remove unnecessary if in qc_pkt_may_rm_hp() Remove the shortcut to use the INITIAL encryption level when removing header protection on first connection packet. This change is useful for the following change which removes ssl_sock_ctx in argument lists in favor of the quic_conn instance.	2021-12-23 16:02:24 +01:00
Amaury Denoyelle	7ca7c84fb8	MINOR: quic: store ssl_sock_ctx reference into quic_conn Add a pointer in quic_conn to its related ssl_sock_ctx. This change is required to avoid to use the connection instance to access it. This commit is part of the rearchitecture of xprt-quic layers and the separation between xprt and connection instances. It will be notably useful when the connection allocation will be delayed.	2021-12-23 15:51:00 +01:00
Amaury Denoyelle	a83729e9e6	MINOR: quic: remove unnecessary call to free_quic_conn_cids() free_quic_conn_cids() was called in quic_build_post_handshake_frames() if an error occured. However, the only error is an allocation failure of the CID which does not required to call it. This change is required for future refcount implementation. The CID lock will be removed from the free_quic_conn_cids() and to the caller.	2021-12-23 15:51:00 +01:00
Amaury Denoyelle	250ac42754	BUG/MINOR: quic: upgrade rdlock to wrlock for ODCID removal When a quic_conn is found in the DCID tree, it can be removed from the first ODCID tree. However, this operation must absolutely be run under a write-lock to avoid race condition. To avoid to use the lock too frequently, node.leaf_p is checked. This value is set to NULL after ebmb_delete.	2021-12-23 15:51:00 +01:00
Amaury Denoyelle	d6b166787c	REORG: quic: remove qc_ prefix on functions which not used it directly The qc_* prefix should be reserved to functions which used a specific quic_conn instance and are expected to be pinned on the connection thread.	2021-12-23 15:51:00 +01:00
Frédéric Lécaille	010e532e81	MINOR: quic: Add CONNECTION_CLOSE phrase to trace Some applications may send some information about the reason why they decided to close a connection. Add them to CONNECTION_CLOSE frame traces. Take the opportunity of this patch to shorten some too long variable names without any impact.	2021-12-23 15:48:25 +01:00
Frédéric Lécaille	1ede823d6b	MINOR: quic: Add traces for RX frames (flow control related) Add traces about important frame types to chunk_tx_frm_appendf() and call this function for any type of frame when parsing a packet. Move it to quic_frame.c	2021-12-23 15:48:25 +01:00
Willy Tarreau	77bfa66124	DEBUG: ssl: make sure we never change a servername on established connections Since this case was already met previously with commit `655dec81b` ("BUG/MINOR: backend: do not set sni on connection reuse"), let's make sure that we don't change reused connection settings. This could be generalized to most settings that are only in effect before the handshake in fact (like set_alpn and a few other ones).	2021-12-23 15:44:06 +01:00
Willy Tarreau	0d93a81863	MINOR: pools: work around possibly slow malloc_trim() during gc During 2.4-dev, support for malloc_trim() was implemented to ease release of memory in a stopping process. This was found to be quite effective and later backported to 2.3.7. Then it was found that sometimes malloc_trim() could take a huge time to complete it if was competing with other threads still allocating and releasing memory, reason why it was decided in 2.5-dev to move malloc_trim() under the thread isolation that was already in place in the shared pool version of pool_gc() (this was commit `26ed1835`). However, other instances of pool_gc() that used to call malloc_trim() were not updated since they were not using thread isolation. Currently we have two other such instances, one for when there is absolutely no pool and one for when there are only thread-local pools. Christian Ruppert reported in GH issue #1490 that he's sometimes seeing and old process die upon reload when upgrading from 2.3 to 2.4, and that this happens inside malloc_trim(). The problem is that since 2.4-dev11 with commit `0bae07592` we detect modern libc that provide a faster thread-aware allocator and do not maintain shared pools anymore. As such we're using again the simpler pool_gc() implementations that do not use thread isolation around the malloc_trim() call. All this code was cleaned up recently and the call moved to a new function trim_all_pools(). This patch implements explicit thread isolation inside that function so that callers do not have to care about this anymore. The thread isolation is conditional so that this doesn't affect the one already in place in the larger version of pool_gc(). This way it will solve the problem for all callers. This patch must be backported as far as 2.3. It may possibly require some adaptations. If trim_all_pools() is not present, copy-pasting the tests in each version of pool_gc() will have the same effect. Thanks to Christian for his detailed report and his testing.	2021-12-23 15:44:06 +01:00
Frédéric Lécaille	2c15a66b61	MINOR: quic: Drop asap Retry or Version Negotiation packets These packet are only sent by servers. We drop them as soon as possible when we are an haproxy listener.	2021-12-22 20:43:22 +01:00
Frédéric Lécaille	e7ff2b265a	MINOR: quic: xprt traces fixes Empty parameters are permitted with TRACE_*() macros. If removed, must be replaced by NULL.	2021-12-22 20:43:22 +01:00
Frédéric Lécaille	10250b2e93	MINOR: quic: Handle the cases of overlapping STREAM frames This is the same treatment for bidi and uni STREAM frames. This is a duplication code which should me remove building a function for both these types of streams.	2021-12-22 20:43:22 +01:00
Frédéric Lécaille	01cfec74f5	MINOR: quic: Wrong dropped packet skipping There were cases where some dropped packets were not well skipped. This led the low level QUIC packet parser to continue from wrong packet boundaries.	2021-12-22 20:43:22 +01:00
Frédéric Lécaille	4d118d6a8e	MINOR: quic: unchecked qc_retrieve_conn_from_cid() returned value If qc_retrieve_conn_from_cid() did not manage to retrieve the connection from packet CIDs, we must drop them.	2021-12-22 17:27:51 +01:00
Frédéric Lécaille	677b99dca7	MINOR: quic: Add stream IDs to qcs_push_frame() traces This is only for debug purpose.	2021-12-21 16:06:03 +01:00
Amaury Denoyelle	e770ce3980	MINOR: quic: add quic_conn instance in traces for qc_new_conn The connection instance has been replaced by a quic_conn as first argument to QUIC traces. It is possible to report the quic_conn instance in the qc_new_conn(), contrary to the connection which is not initialized at this stage.	2021-12-21 15:53:19 +01:00
Amaury Denoyelle	7aaeb5b567	MINOR: quic: use quic_conn as argument to traces Replace the connection instance for first argument of trace callback by a quic_conn instance. The QUIC trace module is properly initialized with the first argument refering to a quic_conn. Replace every connection instances in TRACE_* macros invocation in xprt-quic by its related quic_conn. In some case, the connection is still used to access the quic_conn. It may cause some problem on the future when the connection will be completly separated from the xprt layer. This commit is part of the rearchitecture of xprt-quic layers and the separation between xprt and connection instances.	2021-12-21 15:53:19 +01:00
Amaury Denoyelle	baea96400f	MINOR: trace: add quic_conn argument definition Prepare trace support for quic_conn instances as argument. This will be used by the xprt-quic layer in replacement of the connection. This commit is part of the rearchitecture of xprt-quic layers and the separation between xprt and connection instances.	2021-12-21 15:53:19 +01:00
Amaury Denoyelle	4fd53d772f	MINOR: quic: add const qualifier for traces function Add const qualifier on arguments of several dump functions used in the trace callback. This is required to be able to replace the first trace argument by a quic_conn instance. The first argument is a const pointer and so the members accessed through it must also be const.	2021-12-21 15:53:19 +01:00
Amaury Denoyelle	c15dd9214b	MINOR: quic: add reference to quic_conn in ssl context Add a new member in ssl_sock_ctx structure to reference the quic_conn instance if used in the QUIC stack. This member is initialized during qc_conn_init(). This is needed to be able to access to the quic_conn without relying on the connection instance. This commit is part of the rearchitecture of xprt-quic layers and the separation between xprt and connection instances.	2021-12-21 15:53:19 +01:00
Amaury Denoyelle	8a5b27a9b9	REORG: quic: move mux function outside of xprt Move qcc_get_qcs() function from xprt_quic.c to mux_quic.c. This function is used to retrieve the qcs instance from a qcc with a stream id. This clearly belongs to the mux-quic layer.	2021-12-21 15:51:40 +01:00
Amaury Denoyelle	17a741693c	CLEANUP: quic: rename quic_conn instances to qc Use the convention of naming quic_conn instance as qc to not confuse it with a connection instance. The changes occured for qc_parse_pkt_frms(), qc_build_frms() and qc_do_build_pkt().	2021-12-21 15:51:30 +01:00
Frédéric Lécaille	2ce5acf7ed	MINOR: quic: Wrong packet refcount handling in qc_pkt_insert() The QUIC connection I/O handler qc_conn_io_cb() could be called just after qc_pkt_insert() have inserted a packet in a its tree, and before qc_pkt_insert() have incremented the reference counter to this packet. As qc_conn_io_cb() decrement this counter, the packet could be released before qc_pkt_insert() might increment the counter, leading to possible crashes when trying to do so. So, let's make qc_pkt_insert() increment this counter before inserting the packet it is tree. No need to lock anything for that.	2021-12-20 17:33:51 +01:00
Frédéric Lécaille	f1d38cbe15	MINOR: quic: Do not forget STREAM frames received in disorder Add a function to process all STREAM frames received and ordered by their offset (qc_treat_rx_strm_frms()) and modify qc_handle_bidi_strm_frm() consequently.	2021-12-20 17:33:51 +01:00
Frédéric Lécaille	4137b2d316	MINOR: quic: Do not expect to receive only one O-RTT packet There is nothing about this in the RFC. We must support to receive several 0-RTT packets before the handshake has completed.	2021-12-20 17:33:51 +01:00
Frédéric Lécaille	91ac6c3a8a	MINOR: quic: Add a function to list remaining RX packets by encryption level This is only to debug some issues which cause the RX buffer saturation with "Too big packet" traces.	2021-12-20 17:33:51 +01:00
Remi Tricot-Le Breton	cc750efbc5	MINOR: ssl: Remove empty lines from "show ssl ocsp-response" output There were empty lines in the output of the CLI's "show ssl ocsp-response" command (after the certificate ID and between two certificates). This patch removes them since an empty line should mark the end of the output. Must be backported in 2.5.	2021-12-20 12:02:17 +01:00
Amaury Denoyelle	dbef985b74	MINOR: quic: simplify the removal from ODCID tree With the DCID refactoring, the locking is more centralized. It is possible to simplify the code for removal of a quic_conn from the ODCID tree. This operation can be conducted as soon as the connection has been retrieved from the DCID tree, meaning that the peer now uses the final DCID. Remove the bit to flag a connection for removal and just uses ebmb_delete() on each sucessful lookup on the DCID tree. If the quic_conn has already been removed, it is just a noop thanks to eb_delete() implementation.	2021-12-17 10:59:36 +01:00
Amaury Denoyelle	8efe032bba	MINOR: quic: refactor DCID lookup A new function named qc_retrieve_conn_from_cid() now contains all the code to retrieve a connection from a DCID. It handle all type of packets and centralize the locking on the ODCID/DCID trees. This simplify the qc_lstnr_pkt_rcv() function.	2021-12-17 10:59:36 +01:00
Amaury Denoyelle	adb2276524	MINOR: quic: compare coalesced packets by DCID If an UDP datagram contains multiple QUIC packets, they must all use the same DCID. The datagram context is used partly for this. To ensure this, a comparison was made on the dcid_node of DCID tree. As this is a comparison based on pointer address, it can be faulty when nodes are removed/readded on the same pointer address. Replace this comparison by a proper comparison on the DCID data itself. To this end, the dgram_ctx structure contains now a quic_cid member.	2021-12-17 10:59:36 +01:00
Amaury Denoyelle	c92cbfc014	MINOR: quic: refactor concat DCID with address for Initial packets For first Initial packets, the socket source dest address is concatenated to the DCID. This is used to be able to differentiate possible collision between several clients which used the same ODCID. Refactor the code to manage DCID and the concatenation with the address. Before this, the concatenation was done on the quic_cid struct and its <len> field incremented. In the code it is difficult to differentiate a normal DCID with a DCID + address concatenated. A new field <addrlen> has been added in the quic_cid struct. The <len> field now only contains the size of the QUIC DCID. the <addrlen> is first initialized to 0. If the address is concatenated, it will be updated with the size of the concatenated address. This now means we have to explicitely used either cid.len or cid.len + cid.addrlen to access the DCID or the DCID + the address. The code should be clearer thanks to this. The field <odcid_len> in quic_rx_packet struct is now useless and has been removed. However, a new parameter must be added to the qc_new_conn() function to specify the size of the ODCID addrlen.	2021-12-17 10:59:36 +01:00
Amaury Denoyelle	d496251cde	MINOR: quic: rename constant for haproxy CIDs length On haproxy implementation, generated DCID are on 8 bytes, the minimal value allowed by the specification. Rename the constant representing this size to inform that this is haproxy specific.	2021-12-17 10:59:36 +01:00
Amaury Denoyelle	260e5e6c24	MINOR: quic: add missing lock on cid tree All operation on the ODCID/DCID trees must be conducted under a read-write lock. Add a missing read-lock on the lookup operation inside listener handler.	2021-12-17 10:59:36 +01:00
Amaury Denoyelle	67e6cd50ef	CLEANUP: quic: rename quic_conn conn to qc in quic_conn_free Rename quic_conn from conn to qc to differentiate it from a struct connection instance. This convention is already used in the majority of the code.	2021-12-17 10:59:35 +01:00
Amaury Denoyelle	47e1f6d4e2	CLEANUP: quic: fix spelling mistake in a trace Initiial -> Initial	2021-12-17 10:59:35 +01:00
Amaury Denoyelle	fdbf63e86e	MINOR: mux-quic: fix trace on stream creation Replace non-initialized qcs.by_id.key by the id to report the proper stream ID on stream creation.	2021-12-17 09:55:01 +01:00
Frédéric Lécaille	8678eb0d19	CLEANUP: quic: Shorten a litte bit the traces in lstnr_rcv_pkt() Some traces were too long and confusing when displaying 0 for a non-already parsed packet number.	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	25eeebe293	MINOR: quic: Do not mix packet number space and connection flags The packet number space flags were mixed with the connection level flags. This leaded to ACK to be sent at the connection level without regard to the underlying packet number space. But we want to be able to acknowleged packets for a specific packet number space.	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	afd373c232	MINOR: hq_interop: Stop BUG_ON() truncated streams This is required if we do not want to make haproxy crash during zerortt interop runner test which makes a client open multiple streams with long request paths.	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	3fe7df877d	CLEANUP: quic: Comment fix for qc_strm_cpy() This function never returns a negative value... hopefully because it returns a size_t!!!	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	e629cfd96a	MINOR: qpack: Missing check for truncated QPACK fields Decrementing <len> variable without checking could make haproxy crash (on abort) when printing a huge buffer (with negative length).	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	a5da31d186	MINOR: quic: Make xprt support 0-RTT. A client sends a 0-RTT data packet after an Initial one in the same datagram. We must be able to parse such packets just after having parsed the Initial packets.	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	1761fdf0c6	MINOR: ssl_sock: Set the QUIC application from ssl_sock_advertise_alpn_protos. Make this function call quic_set_app_ops() if the protocol could be negotiated by the TLS stack.	2021-12-17 08:38:43 +01:00

... 33 34 35 36 37 ...

16085 Commits