haproxy

mirror of https://git.haproxy.org/git/haproxy.git/ synced 2025-08-15 03:26:59 +02:00

Author	SHA1	Message	Date
Amaury Denoyelle	adb2276524	MINOR: quic: compare coalesced packets by DCID If an UDP datagram contains multiple QUIC packets, they must all use the same DCID. The datagram context is used partly for this. To ensure this, a comparison was made on the dcid_node of DCID tree. As this is a comparison based on pointer address, it can be faulty when nodes are removed/readded on the same pointer address. Replace this comparison by a proper comparison on the DCID data itself. To this end, the dgram_ctx structure contains now a quic_cid member.	2021-12-17 10:59:36 +01:00
Amaury Denoyelle	c92cbfc014	MINOR: quic: refactor concat DCID with address for Initial packets For first Initial packets, the socket source dest address is concatenated to the DCID. This is used to be able to differentiate possible collision between several clients which used the same ODCID. Refactor the code to manage DCID and the concatenation with the address. Before this, the concatenation was done on the quic_cid struct and its <len> field incremented. In the code it is difficult to differentiate a normal DCID with a DCID + address concatenated. A new field <addrlen> has been added in the quic_cid struct. The <len> field now only contains the size of the QUIC DCID. the <addrlen> is first initialized to 0. If the address is concatenated, it will be updated with the size of the concatenated address. This now means we have to explicitely used either cid.len or cid.len + cid.addrlen to access the DCID or the DCID + the address. The code should be clearer thanks to this. The field <odcid_len> in quic_rx_packet struct is now useless and has been removed. However, a new parameter must be added to the qc_new_conn() function to specify the size of the ODCID addrlen.	2021-12-17 10:59:36 +01:00
Amaury Denoyelle	d496251cde	MINOR: quic: rename constant for haproxy CIDs length On haproxy implementation, generated DCID are on 8 bytes, the minimal value allowed by the specification. Rename the constant representing this size to inform that this is haproxy specific.	2021-12-17 10:59:36 +01:00
Frédéric Lécaille	25eeebe293	MINOR: quic: Do not mix packet number space and connection flags The packet number space flags were mixed with the connection level flags. This leaded to ACK to be sent at the connection level without regard to the underlying packet number space. But we want to be able to acknowleged packets for a specific packet number space.	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	a5da31d186	MINOR: quic: Make xprt support 0-RTT. A client sends a 0-RTT data packet after an Initial one in the same datagram. We must be able to parse such packets just after having parsed the Initial packets.	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	b0bd62db23	MINOR: quic: Add quic_set_app_ops() function Export the code responsible which set the ->app_ops structure into quic_set_app_ops() function. It must be called by the TLS callback which selects the application (ssl_sock_advertise_alpn_protos) so that to be able to build application packets after having received 0-RTT data.	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	0371cd54d0	CLEANUP: quic: Remove cdata_len from quic_tx_packet struct This field is no more useful. Modify the traces consequently. Also initialize ->pn_node.key value to -1, which is an illegal value for QUIC packet number, and display it in traces if different from -1.	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	1d2faa24d2	CLEANUP: quic_frame: Remove a useless suffix to STOP_SENDING This is to be consistent with the other frame names. Adding a _frame suffixe to STOP_SENDING is useless. We know this is a frame.	2021-12-17 08:38:43 +01:00
Frédéric Lécaille	f57c333ac1	MINOR: quic: Attach timer task to thread for the connection. This is to avoid races between the connection I/O handler and this task which share too much variables.	2021-12-17 08:38:43 +01:00
Remi Tricot-Le Breton	bb6bc95b1e	MINOR: vars: Parse optional conditions passed to the set-var actions This patch adds the parsing of the optional condition parameters that can be passed to the set-var and set-var-fmt actions (http as well as tcp). Those conditions will not be taken into account yet in the var_set function so conditions passed as parameters will not have any effect. Since actions do not benefit from the parameter preparsing that converters have, parsing conditions needed to be done by hand.	2021-12-16 17:31:57 +01:00
Remi Tricot-Le Breton	51899d251c	MINOR: vars: Parse optional conditions passed to the set-var converter This patch adds the parsing of the optional condition parameters that can be passed to the set-var converter. Those conditions will not be taken into account yet in the var_set function so conditions passed as parameters will not have any effect. This is true for any condition apart from the "ifexists" one that is also used to replace the VF_UPDATEONLY flag that was used to prevent proc scope variable creation from a LUA module.	2021-12-16 17:31:55 +01:00
Daniel Jakots	d1a2e2b0d1	BUILD: ssl: unbreak the build with newer libressl In LibreSSL 3.5.0, BIO is going to become opaque, so haproxy's compat macros will no longer work. The functions they substitute have been available since LibreSSL 2.7.0.	2021-12-15 11:26:31 +01:00
David CARLIER	f5d48f8b3b	MEDIUM: cfgparse: numa detect topology on FreeBSD. allowing for all platforms supporting cpu affinity to have a chance to detect the cpu topology from a given valid node (e.g. DragonflyBSD seems to be NUMA aware from a kernel's perspective and seems to be willing start to provide userland means to get proper info).	2021-12-15 11:05:51 +01:00
Frédéric Lécaille	a842ca1fca	MINOR: quic: Compilation fix for quic_rx_packet_refinc() This was reported by the CI wich clang as compilator. In file included from src/ssl_sock.c:80: include/haproxy/xprt_quic.h:1100:50: error: passing 'int ' to parameter of type 'unsigned int ' converts between pointers to integer types with different sign [-Werror,-Wpointer-sign] } while (refcnt && !HA_ATOMIC_CAS(&pkt->refcnt, &refcnt, refcnt - 1)); ^~~~~~~	2021-12-08 15:30:02 +01:00
Amaury Denoyelle	f3b0ba7dc9	BUG/MINOR: mux-quic: properly initialize flow control Initialize all flow control members on the qcc instance. Without this, the value are undefined and it may be possible to have errors about reached streams limit.	2021-12-08 15:26:16 +01:00
Amaury Denoyelle	5154e7a252	MINOR: quic: notify the mux on CONNECTION_CLOSE The xprt layer is reponsible to notify the mux of a CONNECTION_CLOSE reception. In this case the flag QC_CF_CC_RECV is positionned on the qcc and the mux tasklet is waken up. One of the notable effect of the QC_CF_CC_RECV is that each qcs will be released even if they have remaining data in their send buffers.	2021-12-08 15:26:16 +01:00
Amaury Denoyelle	2873a31c81	MINOR: mux-quic: do not release qcs if there is remaining data to send A qcs is not freed if there is remaining data in its buffer. In this case, the flag QC_SF_DETACH is positionned. The qcc io handler is responsible to remove the qcs if the QC_SF_DETACH is set and their buffers are empty.	2021-12-08 15:26:16 +01:00
Amaury Denoyelle	9beb387dd1	BUILD: mux-quic: fix compilation with DEBUG_MEM_STATS Replace bug.h by api.h in mux_quic header. This is required because bug.h uses atomic operations when compiled with DEBUG_MEM_STATS. api.h takes care of including atomic.h before bug.h.	2021-12-07 17:29:39 +01:00
Amaury Denoyelle	db44338473	MINOR: quic: add HTX EOM on request end Set the HTX EOM flag on RX the app layer. This is required to notify about the end of the request for the stream analyzers, else the request channel never goes to MSG_DONE state.	2021-12-07 17:11:22 +01:00
Frédéric Lécaille	73dcc6ee62	MINOR: quic: Remove QUIC TX packet length evaluation function Remove qc_eval_pkt() which has come with the multithreading support. It was there to evaluate the length of a TX packet before building. We could build from several thread TX packets without consuming a packet number for nothing (when the building failed). But as the TX packet building functions are always executed by the same thread, the one attached to the connection, this does not make sense to continue to use such a function. Furthermore it is buggy since we had to recently pad the TX packet under certain circumstances.	2021-12-07 15:53:56 +01:00
Frédéric Lécaille	fee7ba673f	MINOR: quic: Delete remaining RX handshake packets After the handshake has succeeded, we must delete any remaining Initial or Handshake packets from the RX buffer. This cannot be done depending on the state the connection (->st quic_conn struct member value) as the packet are not received/treated in order.	2021-12-07 15:53:56 +01:00
Frédéric Lécaille	7d807c93f4	MINOR: quic: QUIC encryption level RX packets race issue The tree containing RX packets must be protected from concurrent accesses.	2021-12-07 15:53:56 +01:00
Frédéric Lécaille	d61bc8db59	MINOR: quic: Race issue when consuming RX packets buffer Add a null byte to the end of the RX buffer to notify the consumer there is no more data to treat. Modify quic_rx_packet_pool_purge() which is the function which remove the RX packet from the buffer. Also rename this function to quic_rx_pkts_del(). As the RX packets may be accessed by the QUIC connection handler (quic_conn_io_cb()) the function responsible of decrementing their reference counters must not access other information than these reference counters! It was a very bad idea to try to purge the RX buffer asap when executing this function.	2021-12-07 15:53:56 +01:00
Amaury Denoyelle	84ea8dcbc4	MEDIUM: mux-quic: handle when sending buffer is full Handle the case when the app layer sending buffer is full. A new flag QC_SF_BLK_MROOM is set in this case and the transfer is interrupted. It is expected that then the conn-stream layer will subscribe to SEND. The MROOM flag is reset each time the muxer transfer data from the app layer to its own buffer. If the app layer has been subscribed on SEND it is woken up.	2021-12-07 15:44:45 +01:00
Amaury Denoyelle	a3f222dc1e	MINOR: mux-quic: implement subscribe on stream Implement the subscription in the mux on the qcs instance. Subscribe is now used by the h3 layer when receiving an incomplete frame on the H3 control stream. It is also used when attaching the remote uni-directional streams on the h3 layer. In the qc_send, the mux wakes up the qcs for each new transfer executed. This is done via the method qcs_notify_send(). The xprt wakes up the qcs when receiving data on unidirectional streams. This is done via the method qcs_notify_recv().	2021-12-07 15:44:45 +01:00
Amaury Denoyelle	c2025c1ec6	MEDIUM: quic: detect the stream FIN Set the QC_SF_FIN_STREAM on the app layers (h3 / hq-interop) when reaching the HTX EOM. This is used to warn the mux layer to set the FIN on the QUIC stream.	2021-12-07 15:44:45 +01:00
Amaury Denoyelle	deed777766	MAJOR: mux-quic: implement a simplified mux version Re-implement the QUIC mux. It will reuse the mechanics from the previous mux without all untested/unsupported features. This should ease the maintenance. Note that a lot of features are broken for the moment. They will be re-implemented on the following commits to have a clean commit history.	2021-12-07 15:44:45 +01:00
Christopher Faulet	d6ae912b04	BUILD: bug: Fix error when compiling with -DDEBUG_STRICT_NOCRASH ha_backtrace_to_stderr() must be declared in CRASH_NOW() macro whe HAProxy is compiled with DEBUG_STRICT_NOCRASH. Otherwise an error is reported during compilation: include/haproxy/bug.h:58:26: error: implicit declaration of function ‘ha_backtrace_to_stderr’ [-Werror=implicit-function-declaration] 58 \| #define CRASH_NOW() do { ha_backtrace_to_stderr(); } while (0) This patch must be backported as far as 2.4.	2021-12-03 08:58:28 +01:00
Christopher Faulet	02c893332b	BUG/MEDIUM: h1: Properly reset h1m flags when headers parsing is restarted If H1 headers are not fully received at once, the parsing is restarted a last time when all headers are finally received. When this happens, the h1m flags are sanitized to remove all value set during parsing. But some flags where erroneously preserved. Among others, H1_MF_TE_CHUNKED flag was not removed, what could lead to parsing error. To fix the bug and make things easy, a mask has been added with all flags that must be preserved. It will be more stable. This mask is used to sanitize h1m flags. This patch should fix the issue #1469. It must be backported to 2.5.	2021-12-02 09:46:29 +01:00
Christopher Faulet	c1699f8c1b	MEDIUM: resolvers: No longer store query items in a list into the response When the response is parsed, query items are stored in a list, attached to the parsed response (resolve_response). First, there is one and only one query sent at a time. Thus, there is no reason to use a list. There is a test to be sure there is only one query item in the response. Then, the reference on this query item is only used to validate the domain name is the one requested. So the query list can be removed. We only expect one query item, no reason to loop on query records. In addition, the query domain name is now immediately checked against the resolution domain name. This way, the query item is only manipulated during the response parsing.	2021-12-01 15:21:56 +01:00
Christopher Faulet	4ab2679689	BUG/MINOR: server: Don't rely on last default-server to init server SSL context During post-parsing stage, the SSL context of a server is initialized if SSL is configured on the server or its default-server. It is required to be able to enable SSL at runtime. However a regression was introduced, because the last parsed default-server is used. But it is not necessarily the default-server line used to configure the server. This may lead to erroneously initialize the SSL context for a server without SSL parameter or the skip it while it should be done. The problem is the default-server used to configure a server is not saved during configuration parsing. So, the information is lost during the post-parsing. To fix the bug, the SRV_F_DEFSRV_USE_SSL flag is introduced. It is used to know when a server was initialized with a default-server using SSL. For the record, the commit `f63704488e` ("MEDIUM: cli/ssl: configure ssl on server at runtime") has introduced the bug. This patch must be backported as far as 2.4.	2021-12-01 11:47:08 +01:00
David CARLIER	b1e190a885	MEDIUM: pool: Following up on previous pool trimming update. Apple libmalloc has its own notion of memory arenas as malloc_zone with rich API having various callbacks for various allocations strategies but here we just use the defaults. In trim_all_pools, we advise to purge each zone as much as possible, called "greedy" mode.	2021-12-01 10:38:31 +01:00
Frédéric Lécaille	008386bec4	MINOR: quic: Delete the ODCIDs asap As soon as the connection ID (the one choosen by the QUIC server) has been used by the client, we can delete its original destination connection ID from its tree.	2021-11-30 12:01:32 +01:00
Frédéric Lécaille	40df78f116	MINOR: quic: Add structures to maintain key phase information When running Key Update process, we must maintain much information especially when the key phase bit has been toggled by the peer as it is possible that it is due to late packets. This patch adds quic_tls_kp new structure to do so. They are used to store previous and next secrets, keys and IVs associated to the previous and next RX key phase. We also need the next TX key phase information to be able to encrypt packets for the next key phase.	2021-11-30 11:51:12 +01:00
Frédéric Lécaille	39484de813	MINOR: quic: Add a function to derive the key update secrets This is the function used to derive an n+1th secret from the nth one as described in RFC9001 par. 6.1.	2021-11-30 11:51:12 +01:00
Frédéric Lécaille	fc768ecc88	MINOR: quic: Dynamically allocate the secrete keys This is done for any encryption level. This is to prepare the Key Update feature.	2021-11-30 11:51:12 +01:00
Frédéric Lécaille	1fc5e16c4c	MINOR: quic: More accurate immediately close. When sending a CONNECTION_CLOSE frame to immediately close the connection, do not provide CRYPTO data to the TLS stack. Do not built anything else than a CONNECTION_CLOSE and do not derive any secret when in immediately close state. Seize the opportunity of this patch to rename ->err quic_conn struct member to ->error_code.	2021-11-30 11:47:46 +01:00
Frédéric Lécaille	067a82bba1	MINOR: quic: Set "no_application_protocol" alert We set this TLS error when no application protocol could be negotiated via the TLS callback concerned. It is converted as a QUIC CRYPTO_ERROR error (0x178).	2021-11-30 11:47:46 +01:00
Amaury Denoyelle	d6a352a58b	MEDIUM: quic: handle CIDs to rattach received packets to connection Change the way the CIDs are organized to rattach received packets DCID to QUIC connection. This is necessary to be able to handle multiple DCID to one connection. For this, the quic_connection_id structure has been extended. When allocated, they are inserted in the receiver CID tree instead of the quic_conn directly. When receiving a packet, the receiver tree is inspected to retrieve the quic_connection_id. The quic_connection_id contains now contains a reference to the QUIC connection.	2021-11-25 11:41:29 +01:00
Amaury Denoyelle	42b9f1c6dd	CLEANUP: quic: add comments on CID code Add minor comment to explain how the CID are stored in the QUIC connection.	2021-11-25 11:33:35 +01:00
Willy Tarreau	73dec76e85	[RELEASE] Released version 2.6-dev0 Released version 2.6-dev0 with the following main changes : - MINOR: version: it's development again	2021-11-23 15:50:11 +01:00
Willy Tarreau	3b068c45ee	MINOR: version: it's development again This essentially reverts `9dc4057df0`.	2021-11-23 15:48:35 +01:00
Willy Tarreau	9dc4057df0	MINOR: version: mention that it's stable now This version will be maintained up to around Q1 2023. The INSTALL file also mentions it.	2021-11-23 15:38:10 +01:00
Ilya Shipitsin	a4d09e7ffd	CLEANUP: assorted typo fixes in the code and comments This is 28th iteration of typo fixes	2021-11-22 19:08:12 +01:00
Frédéric Lécaille	66cbb8232c	MINOR: quic: Send CONNECTION_CLOSE frame upon TLS alert Add ->err member to quic_conn struct to store the connection errors. This is the responsability of ->send_alert callback of SSL_QUIC_METHOD struct to handle the TLS alert and consequently update ->err value. At this time, when entering qc_build_pkt() we build a CONNECTION_CLOSE frame close the connection when ->err value is not null.	2021-11-19 14:37:35 +01:00
Frédéric Lécaille	2b2032adbb	MINOR: quic: Update some QUIC protocol errors Add QC_ERR_ prefix to the macro names for the protocol errors. Add new ones which came with the last RFC drafts.	2021-11-19 14:37:35 +01:00
Frédéric Lécaille	ca98a7f9c0	MINOR: quic: Anti-amplification implementation A QUIC server MUST not send more than three times as many bytes as received by clients before its address validation.	2021-11-19 14:37:35 +01:00
Frédéric Lécaille	a956d15118	MINOR: quic: Support transport parameters draft TLS extension If we want to run quic-tracker against haproxy, we must at least support the draft version of the TLS extension for the QUIC transport parameters (0xffa5). quic-tracker QUIC version is draft-29 at this time. We select this depending on the QUIC version. If draft, we select the draft TLS extension.	2021-11-19 14:37:35 +01:00
William Lallemand	e18d4e8286	BUG/MEDIUM: ssl: backend TLS resumption with sni and TLSv1.3 When establishing an outboud connection, haproxy checks if the cached TLS session has the same SNI as the connection we are trying to resume. This test was done by calling SSL_get_servername() which in TLSv1.2 returned the SNI. With TLSv1.3 this is not the case anymore and this function returns NULL, which invalidates any outboud connection we are trying to resume if it uses the sni keyword on its server line. This patch fixes the problem by storing the SNI in the "reused_sess" structure beside the session itself. The ssl_sock_set_servername() now has a RWLOCK because this session cache entry could be accessed by the CLI when trying to update a certificate on the backend. This fix must be backported in every maintained version, however the RWLOCK only exists since version 2.4.	2021-11-19 03:58:30 +01:00
Amaury Denoyelle	154bc7f864	MINOR: quic: support hq-interop Implement a new app_ops layer for quic interop. This layer uses HTTP/0.9 on top of QUIC. Implementation is minimal, with the intent to be able to pass interoperability test suite from https://github.com/marten-seemann/quic-interop-runner. It is instantiated if the negotiated ALPN is "hq-interop".	2021-11-18 10:50:58 +01:00
Amaury Denoyelle	71e588c8a7	MEDIUM: quic: inspect ALPN to install app_ops Remove the hardcoded initialization of h3 layer on mux init. Now the ALPN is looked just after the SSL handshake. The app layer is then installed if the ALPN negotiation returned a supported protocol. This required to add a get_alpn on the ssl_quic layer which is just a call to ssl_sock_get_alpn() from ssl_sock. This is mandatory to be able to use conn_get_alpn().	2021-11-18 10:50:58 +01:00
Amaury Denoyelle	abbe91e5e8	MINOR: quic: redirect app_ops snd_buf through mux This change is required to be able to use multiple app_ops layer on top of QUIC. The stream-interface will now call the mux snd_buf which is just a proxy to the app_ops snd_buf function. The architecture may be simplified in the structure to install the app_ops on the stream_interface and avoid the detour via the mux layer on the sending path.	2021-11-18 10:50:58 +01:00
Willy Tarreau	8c13a92da0	BUG/MEDIUM: connection: make cs_shutr/cs_shutw//cs_close() idempotent In 1.8 when muxes and conn_streams were introduced, the call to conn_full_close() was replaced with a call to cs_close() which only relied on shutr/shutw (commits `6978db35e` ("MINOR: connection: add cs_close() to close a conn_stream") and `a553ae96f` ("MEDIUM: connection: replace conn_full_close() with cs_close()")). By then this was fine, and the rare risk of non-idempotent calls was addressed by the muxes implementing the functions (e.g. mux_pt). Later with commit `325607397` ("MEDIUM: stream: do not forcefully close the client connection anymore"), stream_free() started to call cs_close() instead of forcibly closing the connection via conn_full_close(). At this point this started to break idempotence because it was possible to emit a shutw() (e.g. when option httpclose was set), then to have it called agian upon stream_free() via cs_close(). By then it was not a problem since only mux_pt would implement this and did check for idempotence. When HTX was implemented and mux-h1/h2 offered support for shutw/shutr, the idempotence changed a little bit because the last shutdown mode (normal/silent) was recorded and used at the moment of closing. The call to cs_close() uses the silent mode and will replace the current one. This has an effect on data pending in the buffer if the FIN could not be sent before cs_close(), because lingering may be disabled and final data lost in the network stack. Interestingly, during 2.4-dev3, this was addressed as the side effect of an improvement by commit `3c82d8b32` ("MINOR: mux-h1: Rework how shutdowns are handled"), where the H1 mux's shutdown function becomes explicitly idempotent. However older versions (2.3 to 2.0) do not have it. This patch addresses the issue globally by making sure that cs_shutr() and cs_shutw() are idempotent and cannot have their data truncated by a late cs_close(). It fixes the truncation that is observed in 2.3 and 2.2 as described in issue #1450. This must be backported as far as 2.0, along with commit `f14d750bf` ("BUG/MEDIUM: conn-stream: Don't reset CS flags on close") which it depends on.	2021-11-14 13:42:17 +01:00
William Lallemand	6883674084	MINOR: mworker: implement a reload failure counter Implement a reload failure counter which counts the number of failure since the last success. This counter is available in 'show proc' over the master CLI.	2021-11-10 15:53:01 +01:00
Christopher Faulet	f14d750bf7	BUG/MEDIUM: conn-stream: Don't reset CS flags on close cs_close() and cs_drain_and_close() are called to close a conn-stream. cs_shutr() and cs_shutw() are called with the appropriate modes. But the conn-stream is not released at this stage. However the flags are reset. Thus, after a cs_close(), we loose shutdown flags. If cs_close() is performed several times, it is a problem. And indeed, it is possible. On one hand, the stream-interface may close the conn-stream. On the other end, the stream always closes it when it is released. It is a problem for the H1 multiplexer. Because the conn-stream flags are reset, the shutr and shutw are performed twice. For a delayed shutdown, the dirty mode is used instead of the normal one because the last call to h1_shutw() overwrite H1C flags set by the first call. This leads to dirty shutdowns while normal ones are required. At the end, it is possible to truncate the messages. This bug was revealed by the commit `a85c522d4` ("BUG/MINOR: mux-h1: Save shutdown mode if the shutdown is delayed"). This patch is related to the issue #1450. It must be backported as far as 2.0.	2021-11-10 15:12:49 +01:00
William Dauchy	42d7c402d5	MINOR: promex: backend aggregated server check status - add new metric: `haproxy_backend_agg_server_check_status` it counts the number of servers matching a specific check status this permits to exclude per server check status as the usage is often to rely on the total. Indeed in large setup having thousands of servers per backend the memory impact is not neglible to store the per server metric. - realign promex_str_metrics array quite simple implementation - we could improve it later by adding an internal state to the prometheus exporter, thus to avoid counting at every dump. this patch is an attempt to close github issue #1312. It may bebackported to 2.4 if requested. Signed-off-by: William Dauchy <wdauchy@gmail.com>	2021-11-09 10:51:08 +01:00
William Dauchy	d3141b1d37	DOC: stats: fix location of the text representation `info_field_names` and `stat_field_names` no longer exist and have been moved in stats.c To avoid changing this comment, just mention the name of the new table `info_fields` and `stat_fields` Signed-off-by: William Dauchy <wdauchy@gmail.com>	2021-11-08 13:46:02 +01:00
Willy Tarreau	49b0482ed4	CLEANUP: chunk: remove misleading chunk_strncat() function This function claims to perform an strncat()-like operation but it does not, it always copies the indicated number of bytes, regardless of the presence of a NUL character (what is currently done by chunk_memcat()). Let's remove it and explicitly replace it with chunk_memcat().	2021-11-08 12:08:26 +01:00
Tim Duesterhus	8f1d178cd1	CLEANUP: chunk: Remove duplicated chunk_Xcat implementation Delegate chunk_istcat, chunk_cat and chunk_strncat to the most generic chunk_memcat.	2021-11-08 12:08:26 +01:00
Willy Tarreau	6f7497616e	MEDIUM: connection: rename fc_conn_err and bc_conn_err to fc_err and bc_err Commit `3d2093af9` ("MINOR: connection: Add a connection error code sample fetch") added these convenient sample-fetch functions but it appears that due to a misunderstanding the redundant "conn" part was kept in their name, causing confusion, since "fc" already stands for "front connection". Let's simply call them "fc_err" and "bc_err" to match all other related ones before they appear in a final release. The VTC they appeared in were also updated, and the alpha sort in the keywords table updated. Cc: William Lallemand <wlallemand@haproxy.org>	2021-11-06 09:20:07 +01:00
Frédéric Lécaille	46ea033be0	MINOR: quic: Remove a useless lock for CRYPTO frames ->frms_rwlock is an old lock supposed to be used when several threads could handle the same connection. This is no more the case since this commit: "MINOR: quic: Attach the QUIC connection to a thread."	2021-11-05 15:20:04 +01:00
Frédéric Lécaille	324ecdafbb	MINOR: quic: Enhance the listener RX buffering part Add a buffer per QUIC connection. At this time the listener which receives the UDP datagram is responsible of identifying the underlying QUIC connection and must copy the QUIC packets to its buffer. ->pkt_list member has been added to quic_conn struct to enlist the packets in the order they have been copied to the connection buffer so that to be able to consume this buffer when the packets are freed. This list is locked thanks to a R/W lock to protect it from concurent accesses. quic_rx_packet struct does not use a static buffer anymore to store the QUIC packets contents.	2021-11-05 15:20:04 +01:00
Frédéric Lécaille	c5c69a0ad2	CLEANUP: quic: Remove useless code Remove old I/O handler implementation (listener and server). At this time keep a defined but not used function for servers (qc_srv_pkt_rcv()).	2021-11-05 15:20:04 +01:00
Frédéric Lécaille	c1029f6182	MINOR: quic: Allocate listener RX buffers At this time we allocate an RX buffer by thread. Also take the opportunity offered by this patch to rename TX related variable names to distinguish them from the RX part.	2021-11-05 15:20:04 +01:00
Emeric Brun	f8642ee826	MEDIUM: resolvers: rename dns extra counters to resolvers extra counters This patch renames all dns extra counters and stats functions, types and enums using the 'resolv' prefix/suffixes. The dns extra counter domain id used on cli was replaced by "resolvers" instead of "dns". The typed extra counter prefix dumping resolvers domain "D." was also renamed "N." because it points counters on a Nameserver. This was done to finish the split between "resolver" and "dns" layers and to avoid further misunderstanding when haproxy will handle dns load balancing. This should not be backported.	2021-11-03 17:16:46 +01:00
Emeric Brun	d174f0e59a	MINOR: resolvers/dns: split dns and resolver counters in dns_counter struct This patch add a union and struct into dns_counter struct to split application specific counters. The only current existing application is the resolver.c layer but in futur we could handle different application such as dns load balancing with others specific counters. This patch should not be backported.	2021-11-03 17:16:46 +01:00
Amaury Denoyelle	9c3251d108	MEDIUM: server/backend: implement websocket protocol selection Handle properly websocket streams if the server uses an ALPN with both h1 and h2. Add a new field h2_ws in the server structure. If set to off, reuse is automatically disable on backend and ALPN is forced to http1.x if possible. Nothing is done if on. Implement a mechanism to be able to use a different http version for websocket streams. A new server member <ws> represents the algorithm to select the protocol. This can overrides the server <proto> configuration. If the connection uses ALPN for proto selection, it is updated for websocket streams to select the right protocol. Three mode of selection are implemented : - auto : use the same protocol between non-ws and ws streams. If ALPN is use, try to update it to "http/1.1"; this is only done if the server ALPN contains "http/1.1". - h1 : use http/1.1 - h2 : use http/2.0; this requires the server to support RFC8441 or an error will be returned by haproxy.	2021-11-03 16:24:48 +01:00
Amaury Denoyelle	ac03ef26e8	MINOR: connection: add alternative mux_ops param for conn_install_mux_be Add a new parameter force_mux_ops. This will be useful to specify an alternative to the srv->mux_proto field. If non-NULL, it will be use to force the mux protocol wether srv->mux_proto is set or not. This argument will become useful to install a mux for non-standard streams, most notably websocket streams.	2021-11-03 16:24:48 +01:00
Amaury Denoyelle	2454bda140	MINOR: connection: implement function to update ALPN Implement a new function to update the ALPN on an existing connection. on an existing connection. The ALPN from the ssl context can be checked to update the ALPN only if it is a subset of the context value. This method will be useful to change a connection ALPN for websocket, must notably if the server does not support h2 websocket through the rfc8441 Extended Connect.	2021-11-03 16:24:48 +01:00
Amaury Denoyelle	90ac605ef3	MINOR: stream/mux: implement websocket stream flag Define a new stream flag SF_WEBSOCKET and a new cs flag CS_FL_WEBSOCKET. The conn-stream flag is first set by h1/h2 muxes if the request is a valid websocket upgrade. The flag is then converted to SF_WEBSOCKET on the stream creation. This will be useful to properly manage websocket streams in connect_server().	2021-11-03 16:24:48 +01:00
David Carlier	5592c7b455	BUILD/MINOR: cpuset freebsd build fix Add missing strings.h header. Required for ffsl definition used by CPU_FFS macro. This must be backported up to 2.4.	2021-11-02 13:58:28 +01:00
William Lallemand	bd5739e93e	MINOR: httpclient/lua: handle the streaming into the lua applet With this feature the lua implementation of the httpclient is now able to stream a payload larger than an haproxy buffer. The hlua_httpclient_send() function is now split into: hlua_httpclient_send() which initiate the httpclient and parse the lua parameters hlua_httpclient_snd_yield() which will send the request and be called again to stream the request if the body is larger than an haproxy buffer hlua_httpclient_rcv_yield() which will receive the response and store it in the lua buffer.	2021-10-28 16:24:14 +02:00
William Lallemand	0da616ee18	MINOR: httpclient: request streaming with a callback This patch add a way to handle HTTP requests streaming using a callback. The end of the data must be specified by using the "end" parameter in httpclient_req_xfer().	2021-10-28 16:24:14 +02:00
Willy Tarreau	04065b87ce	MINOR: atomic: remove the memcpy() call and dependency on string.h The memcpy() call in the aarch64 version of __ha_cas_dw() is sometimes inlined and sometimes not, depending on the gcc version. It's only used to copy two void*, so let's use direct assignment instead of memcpy(). It would also be possible to change the asm code to directly write there, but it's not worth it. With this change the code is 8kB smaller with gcc-5.4.	2021-10-28 09:45:48 +02:00
David CARLIER	c050dc6c68	BUILD: atomic: fix build on mac/arm64 The inline assembly is invalid for this platform thus falling back to the builtin instead.	2021-10-28 09:45:48 +02:00
Willy Tarreau	4957a32f9e	BUILD: atomic: prefer __atomic_compare_exchange_n() for __ha_cas_dw() __atomic_compare_exchange() is incorrectly documented in the gcc builtins doc, it says the desired value is "type desired" while in reality it is "const type desired" as expected since that value must in no way be modified by the operation. However it seems that clang has implemented it as documented, and reports build warnings when fed a const. This is quite problematic because it means we have to betry the callers, pretending we won't touch their constants but not knowing what the compiler would do with them, and possibly hiding future bugs. Instead of forcing a cast, let's just switch to the better __atomic_compare_exchange_n() that takes a value instead of a pointer. At least with this one there is no doubt about how the input will be used. It was verified that the output object code is the same both in clang and gcc with this change.	2021-10-28 09:45:48 +02:00
Willy Tarreau	14e7f29e86	MINOR: protocols: replace protocol_by_family() with protocol_lookup() At a few places we were still using protocol_by_family() instead of the richer protocol_lookup(). The former is limited as it enforces SOCK_STREAM and a stream protocol at the control layer. At least with protocol_lookup() we don't have this limitationn. The values were still set for now but later we can imagine making them configurable on the fly.	2021-10-27 17:41:07 +02:00
Willy Tarreau	e3b4518414	MINOR: protocols: make use of the protocol type to select the protocol Instead of using sock_type and ctrl_type to select a protocol, let's make use of the new protocol type. For now they always match so there is no change. This is applied to address parsing and to socket retrieval from older processes.	2021-10-27 17:31:20 +02:00
Willy Tarreau	337edfdbc5	MINOR: protocols: add a new protocol type selector The protocol selection is currently performed based on the family, control type and socket type. But this is often not enough, as both only provide DGRAM or STREAM, leaving few variants. Protocols like SCTP for example might be indistinguishable from TCP here. Same goes for TCP extensions like MPTCP. This commit introduces a new enum proto_type that is placed in each and every protocol definition, that will usually more or less match the sock_type, but being an enum, will support additional values.	2021-10-27 17:05:36 +02:00
Christopher Faulet	16f16afb31	MINOR: stream: Use backend stream-interface dst address instead of target_addr target_addr field in the stream structure is removed. The backend stream-interface destination address is now used.	2021-10-27 11:35:59 +02:00
Christopher Faulet	859ff84f8c	MINOR: stream-int: Add src and dst addresses to the stream-interface For now, these addresses are never set. But the idea is to be able to set, at least first, the client source and destination addresses at the stream level without updating the session or connection ones. Of course, because these addresses are carried by the strream-interface, it would be possible to set server source and destination addresses at this level too. Functions to fill these addresses have been added: si_get_src() and si_get_dst(). If not already set, these functions relies on underlying layers to fill stream-interface addresses. On the frontend side, the session addresses are used if set, otherwise the client connection ones are used. On the backend side, the server connection addresses are used. And just like for sessions and conncetions, si_src() and si_dst() may be used to get source and destination addresses or the stream-interface. And, if not set, same mechanism as above is used.	2021-10-27 11:34:21 +02:00
Christopher Faulet	f46e1ea1ad	MINOR: session: Add src and dst addresses to the session For now, these addresses are never set. But the idea is to be able to set client source and destination addresses at the session level without updating the connection ones. Functions to fill these addresses have been added: sess_get_src() and sess_get_dst(). If not already set, these functions relies on conn_get_src() and conn_get_dst() to fill session addresses. And just like for conncetions, sess_src() and sess_dst() may be used to get source and destination addresses. However, if not set, the corresponding address from the underlying client connection is returned. When this happens, the addresses is filled in the connection object.	2021-10-27 11:34:21 +02:00
Christopher Faulet	cc6fc26bfe	MINOR: connection: Add function to get src/dst without updating the connection conn_get_src() and conn_get_dst() functions are used to fill the source and destination addresses of a connection. On success, ->src and ->dst connection fields can be safely used. For convenience, 2 new functions are added here: conn_src() and conn_dst(). These functions return the corresponding address, as a const and only if it is already set. Otherwise NULL is returned.	2021-10-27 11:34:21 +02:00
Christopher Faulet	99163b75ec	CLEANUP: tools: Use const address for get_net_port() and get_host_port() These functions only extract the port from an address. There is no reason to not use a const address.	2021-10-27 11:34:21 +02:00
Christopher Faulet	4bfce397b8	CLEANUP: connection: No longer export make_proxy_line_v1/v2 functions These functions are only used by the make_proxy_line() function. Thus, we can turn them as static.	2021-10-27 11:34:14 +02:00
Christopher Faulet	24a58fbd7e	CLEANUP: lua: Remove any ambiguities about lua txn execution context flags Flags used to set the execution context of a lua txn are used as an enum. It is not uncommon but there are few flags otherwise. So to remove ambiguities, a comment and a _NONE value are added to have a clear definition of supported values. This patch should fix the issue #1429. No backport needed.	2021-10-27 11:04:16 +02:00
William Lallemand	dec25c3e14	MINOR: httpclient: support payload within a buffer httpclient_req_gen() takes a payload argument which can be use to put a payload in the request. This payload can only fit a request buffer. This payload can also be specified by the "body" named parameter within the lua. httpclient. It is also used within the CLI httpclient when specified as a CLI payload with "<<".	2021-10-27 10:19:41 +02:00
Amaury Denoyelle	28c5b3c0bc	BUILD: fix compilation on NetBSD Use include file <sys/time.h> to fix compilation error with timeval in some files. This is as reported as 'man 7 system_data_types'. The build error is reported on NetBSD 9.2. This should be backported up to 2.2.	2021-10-22 17:04:35 +02:00
Frédéric Lécaille	46be7e92b4	MINOR: quic: Increase the size of handshake RX UDP datagrams Some browsers may send Initial packets with sizes greater than 1252 bytes (QUIC_INITIAL_IPV4_MTU). Let us increase this size limit up to 2048 bytes. Also use this size for "max_udp_payload_size" transport parameter to limit the size of the datagrams we want to receive.	2021-10-22 15:48:19 +02:00
Willy Tarreau	20b622e04b	MINOR: connection: add a new CO_FL_WANT_DRAIN flag to force drain on close Sometimes we'd like to do our best to drain pending data before closing in order to save the peer from risking to receive an RST on close. This adds a new connection flag CO_FL_WANT_DRAIN that is used to trigger a call to conn_ctrl_drain() from conn_ctrl_close(), and the sock_drain() function ignores fd_recv_ready() if this flag is set, in order to catch latest data. It's not used for now.	2021-10-21 21:48:23 +02:00
Willy Tarreau	3193eb9907	BUG/MINOR: task: do not set TASK_F_USR1 for no reason This applicationn specific flag was added in 2.4-dev by commit `6fa8bcdc7` ("MINOR: task: add an application specific flag to the state: TASK_F_USR1") to help preserve a the idle connections status across wakeup calls. While the code to do this was OK for tasklets, it was wrong for tasks, as in an effort not to lose it when setting the RUNNING flag (that tasklets don't have), it ended up being inconditionally set. It just happens that for now no regular tasks use it, only tasklets. This fix makes sure we always atomically perform (state & flags \| running) there, using a CAS. It also does it for tasklets because it was possible to lose some such flags if set by another thread, even though this should not happen with current code. In order to make the code more readable (and avoid the previous mistake of repeated flags in the bit field), a new TASK_PERSISTENT aggregate was declared in task.h for this. In practice the CAS is cheap here because task states are stable or convergent so the loop will almost never be taken. This should be backported to 2.4.	2021-10-21 16:17:29 +02:00
Willy Tarreau	c79f014972	MINOR: list: add new macro LIST_INLIST_ATOMIC() This macro is similar to LIST_INLIST() except that it is guaranteed to perform the test atomically, so that even if LIST_INLIST() is intrumented with debugging code to perform extra consistency checks, it will not fail when used in the context of barriers and atomic ops.	2021-10-21 15:28:24 +02:00
Willy Tarreau	9628c42284	OPTIM: resolvers: move the eb32 node before the data in the answer_item perf top shows that we spend a lot of time trying to read item->type in the lookup loop, because the node is placed after the very long name, so when the node is found, no data is in the cache yet. Let's simply move the node upper in the struct. This results in the CPU usage of resolv_validate_dns_response() to drop by 4 points.	2021-10-21 15:28:24 +02:00
Willy Tarreau	dd362b7b24	BUG/MAJOR: buf: fix varint API post- vs pre- increment A bogus test in b_get_varint(), b_put_varint(), b_peek_varint() shifts the end of the buffer by one byte. Since the bug is the same in the read and write functions, the buffer contents remain compatible, which explains why this bug was not detected earlier. But if the buffer ends on an aligned address or page, it can result in a one-byte overflow which will typically cause a crash or an inconsistent behavior. This API is only used by rings (e.g. for traces and boot messages) and by DNS responses, so the probability to hit it is extremely low, but a crash on boot was observed. This must be backported to 2.2.	2021-10-21 15:28:24 +02:00
Willy Tarreau	7893ae117f	MEDIUM: resolvers: replace the answer_list with a (flat) tree With SRV records, a huge amount of time is spent looking for records by walking long lists. It is possible to reduce this by indexing values in trees instead. However the whole code relies a lot on the list ordering, and even implements some round-robin on it to distribute IP addresses to servers. This patch starts carefully by replacing the list with a an eb32 tree that is still used like a list, with a constant key 0. Since ebtrees preserve insertion order for duplicates, the tree walk visits the nodes in the exact same order it did with the lists. This allows to implement the required infrastructure without changing the behavior.	2021-10-21 08:02:08 +02:00
Willy Tarreau	6878f80427	MEDIUM: resolvers: remove the last occurrences of the "safe" argument This one was used to indicate whether the callee had to follow particularly safe code path when removing resolutions. Since the code now uses a kill list, this is not needed anymore.	2021-10-20 17:54:27 +02:00
Willy Tarreau	2acc160c05	CLEANUP: resolvers: do not export resolv_purge_resolution_answer_records() This code is dangerous enough that we certainly don't want external code to ever approach it, let's not export unnecessary functions like this one. It was made static and a comment was added about its purpose.	2021-10-20 17:52:50 +02:00
Remi Tricot-Le Breton	1c891bcc90	MINOR: jwt: jwt_verify returns negative values in case of error In order for all the error return values to be distributed on the same side (instead of surrounding the success error code), the return values for errors other than a simple verification failure are switched to negative values. This way the result of the jwt_verify converter can be compared strictly to 1 as well relative to 0 (any <= 0 return value is an error). The documentation was also modified to discourage conversion of the return value into a boolean (which would definitely not work).	2021-10-18 16:02:29 +02:00
Tim Duesterhus	f480768d31	CLEANUP: Consistently `unsigned int` for bitfields see `6a0dd73390` Found using GitHub's CodeQL scan.	2021-10-18 09:13:24 +02:00
Ilya Shipitsin	bd6b4be721	CLEANUP: assorted typo fixes in the code and comments This is 27th iteration of typo fixes	2021-10-18 07:26:19 +02:00
Christopher Faulet	56717803e1	MINOR: proxy: Add PR_FL_READY flag on fully configured and usable proxies The PR_FL_READY flags must now be set on a proxy at the end of the configuration validity check to notify it is fully configured and may be safely used. For now there is no real usage of this flag. But it will be usefull for referenced default proxies to finish their configuration only once. This patch is mandatory to support TCP/HTTP rules in defaults sections.	2021-10-15 14:12:19 +02:00
Christopher Faulet	27c8d20451	MINOR: proxy: Be able to reference the defaults section used by a proxy A proxy may now references the defaults section it is used. To do so, a pointer on the default proxy was added in the proxy structure. And a refcount must be used to track proxies using a default proxy. A default proxy is destroyed iff its refcount is equal to zero and when it drops to zero. All this stuff must be performed during init/deinit staged for now. All unreferenced default proxies are removed after the configuration parsing. This patch is mandatory to support TCP/HTTP rules in defaults sections.	2021-10-15 14:12:19 +02:00
Christopher Faulet	b40542000d	MEDIUM: proxy: Warn about ambiguous use of named defaults sections It is now possible to designate the defaults section to use by adding a name of the corresponding defaults section and referencing it in the desired proxy section. However, this introduces an ambiguity. This named defaults section may still be implicitly used by other proxies if it is the last one defined. In this case for instance: default common ... default frt from common ... default bck from common ... frontend fe from frt ... backend be from bck ... listen stats ... Here, it is not really obvious the last section will use the 'bck' defaults section. And it is probably not the expected behaviour. To help users to properly configure their haproxy, a warning is now emitted if a defaults section is explicitly AND implicitly used. The configuration manual was updated accordingly. Because this patch adds a warning, it should probably not be backported to 2.4. However, if is is backported, it depends on commit "MINOR: proxy: Introduce proxy flags to replace disabled bitfield".	2021-10-15 14:12:19 +02:00
Christopher Faulet	dfd10ab5ee	MINOR: proxy: Introduce proxy flags to replace disabled bitfield This change is required to support TCP/HTTP rules in defaults sections. The 'disabled' bitfield in the proxy structure, used to know if a proxy is disabled or stopped, is replaced a generic bitfield named 'flags'. PR_DISABLED and PR_STOPPED flags are renamed to PR_FL_DISABLED and PR_FL_STOPPED respectively. In addition, everywhere there is a test to know if a proxy is disabled or stopped, there is now a bitwise AND operation on PR_FL_DISABLED and/or PR_FL_STOPPED flags.	2021-10-15 14:12:19 +02:00
Willy Tarreau	cc8fd4c040	MINOR: resolvers: merge address and target into a union "data" These two fields are exclusive as they depend on the data type. Let's move them into a union to save some precious bytes. This reduces the struct resolv_answer_item size from 600 to 576 bytes.	2021-10-14 22:52:04 +02:00
Willy Tarreau	b4ca0195a9	BUG/MEDIUM: resolvers: use correct storage for the target address The struct resolv_answer_item contains an address field of type "sockaddr" which is only 16 bytes long, but which is used to store either IPv4 or IPv6. Fortunately, the contents only overlap with the "target" field that follows it and that is large enough to absorb the extra bytes needed to store AAAA records. But this is dangerous as just moving fields around could result in memory corruption. The fix uses a union and removes the casts that were used to hide the problem. Older versions need to be checked and possibly fixed. This needs to be backported anyway.	2021-10-14 22:44:51 +02:00
Willy Tarreau	6dfbef4145	MEDIUM: listener: add the "shards" bind keyword In multi-threaded mode, on operating systems supporting multiple listeners on the same IP:port, this will automatically create this number of multiple identical listeners for the same line, all bound to a fair share of the number of the threads attached to this listener. This can sometimes be useful when using very large thread counts where the in-kernel locking on a single socket starts to cause a significant overhead. In this case the incoming traffic is distributed over multiple sockets and the contention is reduced. Note that doing this can easily increase the CPU usage by making more threads work a little bit. If the number of shards is higher than the number of available threads, it will automatically be trimmed to the number of threads. A special value "by-thread" will automatically assign one shard per thread.	2021-10-14 21:27:48 +02:00
Willy Tarreau	59a877dfd9	MINOR: listeners: add clone_listener() to duplicate listeners at boot time This function's purpose will be to duplicate a listener in INIT state. This will be used to ease declaration of listeners spanning multiple groups, which will thus require multiple FDs hence multiple receivers.	2021-10-14 21:27:48 +02:00
Willy Tarreau	01cac3f721	MEDIUM: listeners: split the thread mask between receiver and bind_conf With groups at some point we'll have to have distinct masks/groups in the receiver and the bind_conf, because a single bind_conf might require to instantiate multiple receivers (one per group). Let's split the thread mask and group to have one for the bind_conf and another one for the receiver while it remains easy to do. This will later allow to use different storage for the bind_conf if needed (e.g. support multiple groups).	2021-10-14 21:27:48 +02:00
William Lallemand	1dbf578ee0	BUILD: jwt: fix declaration of EVP_KEY in jwt-h.h In file included from include/haproxy/jwt.h:25: include/haproxy/jwt-t.h:66:2: error: unknown type name 'EVP_PKEY' EVP_PKEY *pkey; ^ 1 error generated. Fix this compilation issue by inserting openssl-compat.h in jwt-t.h	2021-10-14 17:21:11 +02:00
Remi Tricot-Le Breton	130e142ee2	MEDIUM: jwt: Add jwt_verify converter to verify JWT integrity This new converter takes a JSON Web Token, an algorithm (among the ones specified for JWS tokens in RFC 7518) and a public key or a secret, and it returns a verdict about the signature contained in the token. It does not simply return a boolean because some specific error cases cas be specified by returning an integer instead, such as unmanaged algorithms or invalid tokens. This enables to distinguich malformed tokens from tampered ones, that would be valid format-wise but would have a bad signature. This converter does not perform a full JWT validation as decribed in section 7.2 of RFC 7519. For instance it does not ensure that the header and payload parts of the token are completely valid JSON objects because it would need a complete JSON parser. It only focuses on the signature and checks that it matches the token's contents.	2021-10-14 16:38:14 +02:00
Remi Tricot-Le Breton	864089e0a6	MINOR: jwt: Insert public certificates into dedicated JWT tree A JWT signed with the RSXXX or ESXXX algorithm (RSA or ECDSA) requires a public certificate to be verified and to ensure it is valid. Those certificates must not be read on disk at runtime so we need a caching mechanism into which those certificates will be loaded during init. This is done through a dedicated ebtree that is filled during configuration parsing. The path to the public certificates will need to be explicitely mentioned in the configuration so that certificates can be loaded as early as possible. This tree is different from the ckch one because ckch entries are much bigger than the public certificates used in JWT validation process.	2021-10-14 16:38:12 +02:00
Remi Tricot-Le Breton	e0d3c00086	MINOR: jwt: JWT tokenizing helper function This helper function splits a JWT under Compact Serialization format (dot-separated base64-url encoded strings) into its different sub strings. Since we do not want to manage more than JWS for now, which can only have at most three subparts, any JWT that has strictly more than two dots is considered invalid.	2021-10-14 16:38:10 +02:00
Remi Tricot-Le Breton	7feb361776	MINOR: jwt: Parse JWT alg field The full list of possible algorithms used to create a JWS signature is defined in section 3.1 of RFC7518. This patch adds a helper function that converts the "alg" strings into an enum member.	2021-10-14 16:38:08 +02:00
Remi Tricot-Le Breton	f5dd337b12	MINOR: http: Add http_auth_bearer sample fetch This fetch can be used to retrieve the data contained in an HTTP Authorization header when the Bearer scheme is used. This is used when transmitting JSON Web Tokens for instance.	2021-10-14 16:38:07 +02:00
Amaury Denoyelle	493bb1db10	MINOR: quic: handle CONNECTION_CLOSE frame On receiving CONNECTION_CLOSE frame, the mux is flagged for immediate connection close. A stream is closed even if there is data not ACKed left if CONNECTION_CLOSE has been received.	2021-10-13 16:38:56 +02:00
Amaury Denoyelle	1e308ffc79	MINOR: mux: remove last occurences of qcc ring buffer The mux tx buffers have been rewritten with buffers attached to qcs instances. qc_buf_available and qc_get_buf functions are updated to manipulates qcs. All occurences of the unused qcc ring buffer are removed to ease the code maintenance.	2021-10-13 16:38:56 +02:00
Amaury Denoyelle	cae0791942	MEDIUM: mux-quic: defer stream shut if remaining tx data Defer the shutting of a qcs if there is still data in its tx buffers. In this case, the conn_stream is closed but the qcs is kept with a new flag QC_SF_DETACH. On ACK reception, the xprt wake up the shut_tl tasklet if the stream is flagged with QC_SF_DETACH. This tasklet is responsible to free the qcs and possibly the qcc when all bidirectional streams are removed.	2021-10-13 16:38:56 +02:00
Amaury Denoyelle	d3d97c6ae7	MEDIUM: mux-quic: rationalize tx buffers between qcc/qcs Remove the tx mux ring buffers in qcs, which should be in the qcc. For the moment, use a simple architecture with 2 simple tx buffers in the qcs. The first buffer is used by the h3 layer to prepare the data. The mux send operation transfer these into the 2nd buffer named xprt_buf. This buffer is only freed when an ACK has been received. This architecture is functional but not optimal for two reasons : - it won't limit the buffer usage by connection - each transfer on a new stream requires an allocation	2021-10-13 16:38:56 +02:00
Remi Tricot-Le Breton	b01179aa92	MINOR: ssl: Add ssllib_name_startswith precondition This new ssllib_name_startswith precondition check can be used to distinguish application linked with OpenSSL from the ones linked with other SSL libraries (LibreSSL or BoringSSL namely). This check takes a string as input and returns 1 when the SSL library's name starts with the given string. It is based on the OpenSSL_version function which returns the same output as the "openssl version" command.	2021-10-13 11:28:08 +02:00
Willy Tarreau	c9e4868510	MINOR: rules: add a file name and line number to act_rules These ones are passed on rule creation for the sole purpose of being reported in "show sess", which is not done yet. For now the entries are allocated upon rule creation and freed in free_act_rules().	2021-10-12 07:38:30 +02:00
Willy Tarreau	d535f807bb	MINOR: rules: add a new function new_act_rule() to allocate act_rules Rules are currently allocated using calloc() by their caller, which does not make it very convenient to pass more information such as the file name and line number. This patch introduces new_act_rule() which performs the malloc() and already takes in argument the ruleset (ACT_F_*), the file name and the line number. This saves the caller from having to assing ->from, and will allow to improve the internal storage with more info.	2021-10-12 07:38:30 +02:00
Olivier Houchard	e972c0acde	MINOR: initcall: Rename __GLOBL and __GLOBL1. Rename __GLOBL and __GLOBL1 to __HA_GLOBL and __HA_GLOBL1, as the former are already defined on FreeBSD. This should be backported to 2.4, 2.3 and 2.2.	2021-10-11 00:55:26 +02:00
Willy Tarreau	db2ab8218c	MEDIUM: stick-table: never learn the "conn_cur" value from peers There have been a large number of issues reported with conn_cur synchronization because the concept is wrong. In an active-passive setup, pushing the local connections count from the active node to the passive one will result in the passive node to have a higher counter than the real number of connections. Due to this, after a switchover, it will never be able to close enough connections to go down to zero. The same commonly happens on reloads since the new process preloads its values from the old process, and if no connection happens for a key after the value is learned, it is impossible to reset the previous ones. In active-active setups it's a bit different, as the number of connections reflects the number on the peer that pushed last. This patch solves this by marking the "conn_cur" local and preventing it from being learned from peers. It is still pushed, however, so that any monitoring system that collects values from the peers will still see it. The patch is tiny and trivially backportable. While a change of behavior in stable branches is never welcome, it remains possible to fix issues if reports become frequent.	2021-10-08 17:53:12 +02:00
Willy Tarreau	627def9e50	MINOR: threads: add a new function to resolve config groups and masks In the configuration sometimes we'll omit a thread group number to designate a global thread number range, and sometimes we'll mention the group and designate IDs within that group. The operation is more complex than it seems due to the need to check for ranges spanning between multiple groups and determining groups from threads from bit masks and remapping bit masks between local/global. This patch adds a function to perform this operation, it takes a group and mask on input and updates them on output. It's designed to be used by "bind" lines but will likely be usable at other places if needed. For situations where specified threads do not exist in the group, we have the choice in the code between silently fixing the thread set or failing with a message. For now the better option seems to return an error, but if it turns out to be an issue we can easily change that in the future. Note that it should only happen with "x/even" when group x only has one thread.	2021-10-08 17:22:26 +02:00
Willy Tarreau	d57b9ff7af	MEDIUM: listeners: support the definition of thread groups on bind lines This extends the "thread" statement of bind lines to support an optional thread group number. When unspecified (0) it's an absolute thread range, and when specified it's one relative to the thread group. Masks are still used so no more than 64 threads may be specified at once, and a single group is possible. The directive is not used for now.	2021-10-08 17:22:26 +02:00
Willy Tarreau	b90935c908	MINOR: threads: add the current group ID in thread-local "tgid" variable This is the equivalent of "tid" for ease of access. In the future if we make th_cfg a pure thread-local array (not a pointer), it may make sense to move it there.	2021-10-08 17:22:26 +02:00
Willy Tarreau	43ab05b3da	MEDIUM: threads: replace ha_set_tid() with ha_set_thread() ha_set_tid() was randomly used either to explicitly set thread 0 or to set any possibly incomplete thread during boot. Let's replace it with a pointer to a valid thread or NULL for any thread. This allows us to check that the designated threads are always valid, and to ignore the thread 0's mapping when setting it to NULL, and always use group 0 with it during boot. The initialization code is also cleaner, as we don't pass ugly casts of a thread ID to a pointer anymore.	2021-10-08 17:22:26 +02:00
Willy Tarreau	cc7a11ee3b	MINOR: threads: set the tid, ltid and their bit in thread_cfg This will be a convenient way to communicate the thread ID and its local ID in the group, as well as their respective bits when creating the threads or when only a pointer is given.	2021-10-08 17:22:26 +02:00
Willy Tarreau	6eee85f887	MINOR: threads: set the group ID and its bit in the thread group This will ease the reporting of the current thread group ID when coming from the thread itself, especially since it returns the visible ID, starting at 1.	2021-10-08 17:22:26 +02:00
Willy Tarreau	e6806ebecc	MEDIUM: threads: automatically assign threads to groups This takes care of unassigned threads groups and places unassigned threads there, in a more or less balanced way. Too sparse allocations may still fail though. For now with a maximum group number fixed to 1 nothing can really fail.	2021-10-08 17:22:26 +02:00
Willy Tarreau	fc69e410e6	MINOR: threads: make tg point to the current thread's group A the "tg" thread-local variable now always points to the current thread group. It's pre-initializd to the first one during boot and is set to point to the thread's one by ha_set_tid(). This last one takes care of checking whether the thread group was assigned or not because it may be called during boot before threads are initialized.	2021-10-08 17:22:26 +02:00
Willy Tarreau	d04bc3ac21	MINOR: global: add a new "thread-group" directive This registers a mapping of threads to groups by enumerating for each thread what group it belongs to, and marking the group as assigned. It takes care of checking for redefinitions, overlaps, and holes. It supports both individual numbers and ranges. The thread group is referenced from the thread config.	2021-10-08 17:22:26 +02:00
Willy Tarreau	c33b969e35	MINOR: global: add a new "thread-groups" directive This is used to configure the number of thread groups. For now it can only be 1.	2021-10-08 17:22:26 +02:00
Willy Tarreau	f9662848f2	MINOR: threads: introduce a minimalistic notion of thread-group This creates a struct tgroup_info which knows the thread ID of the first thread in a group, and the number of threads in it. For now there's only one thread group supported in the configuration, but it may be forced to other values for development purposes by defining MAX_TGROUPS, and it's enabled even when threads are disabled and will need to remain accessible during boot to keep a simple enough internal API. For the purpose of easing the configurations which do not specify a thread group, we're starting group numbering at 1 so that thread group 0 can be "undefined" (i.e. for "bind" lines or when binding tasks). The goal will be to later move there some global items that must be made per-group.	2021-10-08 17:22:26 +02:00
Willy Tarreau	6036342f58	MINOR: thread: make "ti" a const pointer and clean up thread_info a bit We want to make sure that the current thread_info accessed via "ti" will remain constant, so that we don't accidentally place new variable parts there and so that the compiler knows that info retrieved from there is not expected to have changed between two function calls. Only a few init locations had to be adjusted to use the array and the rest is unaffected.	2021-10-08 17:22:26 +02:00
Willy Tarreau	b4e34766a3	REORG: thread/sched: move the last dynamic thread_info to thread_ctx The last 3 fields were 3 list heads that are per-thread, and which are: - the pool's LRU head - the buffer_wq - the streams list head Moving them into thread_ctx completes the removal of dynamic elements from the struct thread_info. Now all these dynamic elements are packed together at a single place for a thread.	2021-10-08 17:22:26 +02:00
Willy Tarreau	a0b99536c8	REORG: thread/sched: move the thread_info flags to the thread_ctx The TI_FL_STUCK flag is manipulated by the watchdog and scheduler and describes the apparent life/death of a thread so it changes all the time and it makes sense to move it to the thread's context for an active thread.	2021-10-08 17:22:26 +02:00
Willy Tarreau	45c38e22bf	REORG: thread/clock: move the clock parts of thread_info to thread_ctx The "thread_info" name was initially chosen to store all info about threads but since we now have a separate per-thread context, there is no point keeping some of its elements in the thread_info struct. As such, this patch moves prev_cpu_time, prev_mono_time and idle_pct to thread_ctx, into the thread context, with the scheduler parts. Instead of accessing them via "ti->" we now access them via "th_ctx->", which makes more sense as they're totally dynamic, and will be required for future evolutions. There's no room problem for now, the structure still has 84 bytes available at the end.	2021-10-08 17:22:26 +02:00
Willy Tarreau	1a9c922b53	REORG: thread/sched: move the task_per_thread stuff to thread_ctx The scheduler contains a lot of stuff that is thread-local and not exclusively tied to the scheduler. Other parts (namely thread_info) contain similar thread-local context that ought to be merged with it but that is even less related to the scheduler. However moving more data into this structure isn't possible since task.h is high level and cannot be included everywhere (e.g. activity) without causing include loops. In the end, it appears that the task_per_thread represents most of the per-thread context defined with generic types and should simply move to tinfo.h so that everyone can use them. The struct was renamed to thread_ctx and the variable "sched" was renamed to "th_ctx". "sched" used to be initialized manually from run_thread_poll_loop(), now it's initialized by ha_set_tid() just like ti, tid, tid_bit. The memset() in init_task() was removed in favor of a bss initialization of the array, so that other subsystems can put their stuff in this array. Since the tasklet array has TL_CLASSES elements, the TL_* definitions was moved there as well, but it's not a problem. The vast majority of the change in this patch is caused by the renaming of the structures.	2021-10-08 17:22:26 +02:00
Willy Tarreau	6414e4423c	CLEANUP: wdt: do not remap SI_TKILL to SI_LWP, test the values directly We used to remap SI_TKILL to SI_LWP when SI_TKILL was not available (e.g. FreeBSD) but that's ugly and since we need this only in a single switch/case block in wdt.c it's even simpler and cleaner to perform the two tests there, so let's do this.	2021-10-08 17:22:26 +02:00
Willy Tarreau	b474f43816	MINOR: wdt: move wd_timer to wdt.c The watchdog timer had no more reason for being shared with the struct thread_info since the watchdog is the only user now. Let's remove it from the struct and move it to a static array in wdt.c. This removes some ifdefs and the need for the ugly mapping to empty_t that might be subject to a cast to a long when compared to TIMER_INVALID. Now timer_t is not known outside of wdt.c and clock.c anymore.	2021-10-08 17:22:26 +02:00
Willy Tarreau	2169498941	MINOR: clock: move the clock_ids to clock.c This removes the knowledge of clockid_t from anywhere but clock.c, thus eliminating a source of includes burden. The unused clock_id field was removed from thread_info, and the definition setting of clockid_t was removed from compat.h. The most visible change is that the function now_cpu_time_thread() now takes the thread number instead of a tinfo pointer.	2021-10-08 17:22:26 +02:00
Willy Tarreau	6cb0c391e7	REORG: clock/wdt: move wdt timer initialization to clock.c The code that deals with timer creation for the WDT was moved to clock.c and is called with the few relevant arguments. This removes the need for awareness of clock_id from wdt.c and as such saves us from having to share it outside. The timer_t is also known only from both ends but not from the public API so that we don't have to create a fake timer_t anymore on systems which do not support it (e.g. macos).	2021-10-08 17:22:26 +02:00
Willy Tarreau	44c58da52f	REORG: clock: move the clock_id initialization to clock.c This was previously open-coded in run_thread_poll_loop(). Now that we have clock.c dedicated to such stuff, let's move the code there so that we don't need to keep such ifdefs nor to depend on the clock_id.	2021-10-08 17:22:26 +02:00
Willy Tarreau	2c6a998727	CLEANUP: clock: stop exporting before_poll and after_poll We don't need to export them anymore so let's make them static.	2021-10-08 17:22:26 +02:00
Willy Tarreau	20adfde9c8	MINOR: activity: get the run_time from the clock updates Instead of fiddling with before_poll and after_poll in activity_count_runtime(), the function is now called by clock_entering_poll() which passes it the number of microseconds spent working. This allows to remove all calls to activity_count_runtime() from the pollers.	2021-10-08 17:22:26 +02:00
Willy Tarreau	f9d5e1079c	REORG: clock: move the updates of cpu/mono time to clock.c The entering_poll/leaving_poll/measure_idle functions that were hard to classify and used to move to various locations have now been placed into clock.c since it's precisely about time-keeping. The functions were renamed to clock_*. The samp_time and idle_time values are now static since there is no reason for them to be read from outside.	2021-10-08 17:22:26 +02:00
Willy Tarreau	5554264f31	REORG: time: move time-keeping code and variables to clock.c There is currently a problem related to time keeping. We're mixing the functions to perform calculations with the os-dependent code needed to retrieve and adjust the local time. This patch extracts from time.{c,h} the parts that are solely dedicated to time keeping. These are the "now" or "before_poll" variables for example, as well as the various now_() functions that make use of gettimeofday() and clock_gettime() to retrieve the current time. The "tv_" functions moved there were also more appropriately renamed to "clock_*". Other parts used to compute stolen time are in other files, they will have to be picked next.	2021-10-08 17:22:26 +02:00
Willy Tarreau	de361ad22e	BUILD: connection: avoid a build warning on FreeBSD with SO_USER_COOKIE It was brough by an unneeded addition of a local variable after a test in commit `f7f53afcf` ("BUILD/MEDIUM: tcp: set-mark setting support for FreeBSD."). No backport needed.	2021-10-08 17:21:48 +02:00

1 2 3 4 5 ...

5700 Commits