haproxy

mirror of https://git.haproxy.org/git/haproxy.git/ synced 2025-08-09 00:27:08 +02:00

Author	SHA1	Message	Date
Willy Tarreau	784b868c97	MEDIUM: quic: move conn->qc into conn->handle It was supposed to be there, and probably was not placed there due to historic limitations in listener_accept(), but now there does not seem to be a remaining valid reason for keeping the quic_conn out of the handle. In addition in new_quic_cli_conn() the handle->fd was incorrectly set to the listener's FD.	2022-04-11 19:33:04 +02:00
Willy Tarreau	de827958a2	MEDIUM: ssl: improve retrieval of ssl_sock_ctx and SSL detection Historically there was a single way to have an SSL transport on a connection, so detecting if the transport layer was SSL and a context was present was sufficient to detect SSL. With QUIC, things have changed because QUIC also relies on SSL, but the context is embedded inside the quic_conn and the transport layer doesn't match expectations outside, making it difficult to detect that SSL is in use over the connection. The approach taken here to improve this consists in adding a new method at the transport layer, get_ssl_sock_ctx(), to retrieve this often needed ssl_sock_ctx, and to use this to detect the presence of SSL. This will even allow some simplifications and cleanups to be made in the SSL code itself, and QUIC will be able to provide one to export its ssl_sock_ctx.	2022-04-11 19:33:04 +02:00
Willy Tarreau	cdf7c8e543	MINOR: quic-sock: provide a pair of get_src/get_dst functions These functions will allow the connection layer to retrieve a quic_conn's source or destination when possible. The quic_conn holds the peer's address but not the local one, and the sockets API doesn't always makes that easy for datagrams. Thus for frontend connection what we're doing here is to retrieve the listener's address when the destination address is desired. Now it finally becomes possible to fetch the source and destination using "src" and "dst", and to pass an incoming connection's endpoints via the proxy protocol.	2022-04-11 19:33:04 +02:00
Willy Tarreau	e151609110	MINOR: protocol: add get_src() and get_dst() at the protocol level Right now the proto_fam descriptor provides a family-specific get_src() and get_dst() pair of calls to retrieve a socket's source or destination address. However this only works for connected mode sockets. QUIC provides its own stream protocol, which relies on a datagram protocol underneath, so the get_src()/get_dst() at that protocol's family will not work, and QUIC would need to provide its own. This patch implements get_src() and get_dst() at the protocol level from a connection, and makes sure that conn_get_src()/conn_get_dst() will automatically use them if defined before falling back to the family's pair of functions.	2022-04-11 19:33:04 +02:00
Willy Tarreau	987c08a5e2	MINOR: connection: rearrange conn_get_src/dst to be a bit more extensible We'll want conn_get_src/dst to support other means of retrieving these respective IP addresses, but the functions as they're designed are a bit too restrictive for now. This patch arranges them to have a default error fallback allowing to test different mechanisms. In addition we now make sure the underlying protocol is of type stream before calling the family's get_src/dst as it makes no sense to do that on dgram sockets for example.	2022-04-11 19:33:04 +02:00
Willy Tarreau	07ecfc5e88	MEDIUM: connection: panic when calling FD-specific functions on FD-less conns Certain functions cannot be called on an FD-less conn because they are normally called as part of the protocol-specific setup/teardown sequence. Better place a few BUG_ON() to make sure none of them is called in other situations. If any of them would trigger in ambiguous conditions, it would always be possible to replace it with an error.	2022-04-11 19:31:47 +02:00
Willy Tarreau	e22267971b	MINOR: connection: skip FD-based syscalls for FD-less connections Some syscalls at the TCP level act directly on the FD. Some of them are used by TCP actions like set-tos, set-mark, silent-drop, others try to retrieve TCP info, get the source or destination address. These ones must not be called with an invalid FD coming from an FD-less connection, so let's add the relevant tests for this. It's worth noting that all these ones already have fall back plans (do nothing, error, or switch to alternate implementation).	2022-04-11 19:31:47 +02:00
Willy Tarreau	83a966d025	MINOR: connection: add conn_fd() to retrieve the FD only when it exists There are plenty of places (particularly in debug code) where we try to dump the connection's FD only when the connection is defined. That's already a pain but now it gets one step further with QUIC because we do not want to dump this FD in this case. conn_fd() checks if the connection exists, is ready and is not fd-less, and returns the FD only in this case, otherwise returns -1. This aims at simplifying most of these conditions.	2022-04-11 19:31:47 +02:00
Willy Tarreau	c78a9698ef	MINOR: connection: add a new flag CO_FL_FDLESS on fd-less connections QUIC connections do not use a file descriptor, instead they use the quic equivalent which is the quic_conn. A number of our historical functions at the connection level continue to unconditionally touch the file descriptor and this may have consequences once QUIC starts to be used. This patch adds a new flag on QUIC connections, CO_FL_FDLESS, to mention that the connection doesn't have a file descriptor, hence the FD-based API must never be used on them. From now on it will be possible to intrument existing functions to panic when this flag is present.	2022-04-11 19:31:47 +02:00
William Lallemand	d7bfbe2333	BUILD: ssl: add USE_ENGINE and disable the openssl engine by default The OpenSSL engine API is deprecated starting with OpenSSL 3.0. In order to have a clean build this feature is now disabled by default. It can be reactivated with USE_ENGINE=1 on the build line.	2022-04-11 18:41:24 +02:00
Remi Tricot-Le Breton	b5d968d9b2	MEDIUM: global: Add a "close-spread-time" option to spread soft-stop on time window The new 'close-spread-time' global option can be used to spread idle and active HTTP connction closing after a SIGUSR1 signal is received. This allows to limit bursts of reconnections when too many idle connections are closed at once. Indeed, without this new mechanism, in case of soft-stop, all the idle connections would be closed at once (after the grace period is over), and all active HTTP connections would be closed by appending a "Connection: close" header to the next response that goes over it (or via a GOAWAY frame in case of HTTP2). This patch adds the support of this new option for HTTP as well as HTTP2 connections. It works differently on active and idle connections. On active connections, instead of sending systematically the GOAWAY frame or adding the 'Connection: close' header like before once the soft-stop has started, a random based on the remainder of the close window is calculated, and depending on its result we could decide to keep the connection alive. The random will be recalculated for any subsequent request/response on this connection so the GOAWAY will still end up being sent, but we might wait a few more round trips. This will ensure that goaways are distributed along a longer time window than before. On idle connections, a random factor is used when determining the expire field of the connection's task, which should naturally spread connection closings on the time window (see h2c_update_timeout). This feature request was described in GitHub issue #1614. This patch should be backported to 2.5. It depends on "BUG/MEDIUM: mux-h2: make use of http-request and keep-alive timeouts" which refactorized the timeout management of HTTP2 connections.	2022-04-08 18:15:21 +02:00
Frédéric Lécaille	8c7927c6dd	MINOR: quic_tls: Make key update use of reusable cipher contexts We modify the key update feature implementation to support reusable cipher contexts as this is done for the other cipher contexts for packet decryption and encryption. To do so we attach a context to the quic_tls_kp struct and initialize it each time the underlying secret key is updated. Same thing when we rotate the secrets keys, we rotate the contexts as the same time.	2022-04-08 15:38:29 +02:00
Frédéric Lécaille	f2f4a4eee5	MINOR: quic_tls: Stop hardcoding cipher IV lengths For QUIC AEAD usage, the number of bytes for the IVs is always 12.	2022-04-08 15:38:29 +02:00
Frédéric Lécaille	f4605748f4	MINOR: quic_tls: Add reusable cipher contexts to QUIC TLS contexts Add ->ctx new member field to quic_tls_secrets struct to store the cipher context for each QUIC TLS context TX/RX parts. Add quic_tls_rx_ctx_init() and quic_tls_tx_ctx_init() functions to initialize these cipher context for RX and TX parts respectively. Make qc_new_isecs() call these two functions to initialize the cipher contexts of the Initial secrets. Same thing for ha_quic_set_encryption_secrets() to initialize the cipher contexts of the subsequent derived secrets (ORTT, Handshake, 1RTT). Modify quic_tls_decrypt() and quic_tls_encrypt() to always use the same cipher context without allocating it each time they are called.	2022-04-08 15:38:29 +02:00
Amaury Denoyelle	b515b0af1d	MEDIUM: quic: report closing state for the MUX Define a new API to notify the MUX from the quic-conn when the connection is about to be closed. This happens in the following cases : - on idle timeout - on CONNECTION_CLOSE emission or reception The MUX wake callback is called on these conditions. The quic-conn QUIC_FL_NOTIFY_CLOSE is set to only report once. On the MUX side, connection flags CO_FL_SOCK_RD_SH\|CO_FL_SOCK_WR_SH are set to interrupt future emission/reception. This patch is the counterpart to "MEDIUM: mux-quic: report CO_FL_ERROR on send". Now the quic-conn is able to report its closing, which may be translated by the MUX into a CO_FL_ERROR on the connection for the upper layer. This allows the MUX to properly react to the QUIC closing mechanism for both idle-timeout and closing/draining states.	2022-04-07 10:37:45 +02:00
Amaury Denoyelle	fe035eca3a	MEDIUM: mux-quic: report errors on conn-streams Complete the error reporting. For each attached streams, if CO_FL_ERROR is set, mark them with CS_FL_ERR_PENDING\|CS_FL_ERROR. This will notify the upper layer to trigger streams detach and release of the MUX. This reporting is implemented in a new function qc_wake_some_streams(), called by qc_wake(). This ensures that a lower-layer error is quickly reported to the individual streams.	2022-04-07 10:37:45 +02:00
Amaury Denoyelle	198d35f9c6	MINOR: mux-quic: define is_active app-ops Add a new app layer operation is_active. This can be used by the MUX to check if the connection can be considered as active or not. This is used inside qcc_is_dead as a first check. For example on HTTP/3, if there is at least one bidir client stream opened the connection is active. This explicitly ignore the uni streams used for control and qpack as they can never be closed during the connection lifetime.	2022-04-07 10:23:10 +02:00
Amaury Denoyelle	06890aaa91	MINOR: mux-quic: adjust timeout to accelerate closing Improve timeout handling on the MUX. When releasing a stream, first check if the connection can be considered as dead and should be freed immediatly. This allows to liberate resources faster when possible. If the connection is still active, ensure there is no attached conn-stream before scheduling the timeout. To do this, add a nb_cs field in the qcc structure.	2022-04-07 10:23:10 +02:00
Amaury Denoyelle	846cc046ae	MINOR: mux-quic: factorize conn-stream attach Provide a new function qc_attach_cs. This must be used by the app layer when a conn-stream can be instantiated. This will simplify future development.	2022-04-07 10:10:23 +02:00
Amaury Denoyelle	6057b4090e	CLEANUP: mux-quic: remove unused QC_CF_CC_RECV This flag was used to notify the MUX about a CONNECTION_CLOSE frame reception. It is now unused on the MUX side and can be removed. A new mechanism to detect quic-conn closing will be soon implemented.	2022-04-07 10:10:23 +02:00
Amaury Denoyelle	db71e3bd09	BUG/MEDIUM: quic: ensure quic-conn survives to the MUX Rationalize the lifetime of the quic-conn regarding with the MUX. The quic-conn must not be freed if the MUX is still allocated. This simplify the MUX code when accessing the quic-conn and removed possible segfaults. To implement this, if the quic-conn timer expired, the quic-conn is released only if the MUX is not allocated. Else, the quic-conn is flagged with QUIC_FL_CONN_EXP_TIMER. The MUX is then responsible to call quic_close() which will free the flagged quic-conn.	2022-04-07 10:10:22 +02:00
Frédéric Lécaille	59bf255806	MINOR: quic: Add closing connection state New received packets after sending CONNECTION_CLOSE frame trigger a new CONNECTION_CLOSE frame to be sent. Each time such a frame is sent we increase the number of packet required to send another CONNECTION_CLOSE frame. Rearm only one time the idle timer when sending a CONNECTION_CLOSE frame.	2022-04-06 15:52:35 +02:00
Frédéric Lécaille	47756809fb	MINOR: quic: Add draining connection state. As soon as we receive a CONNECTION_CLOSE frame, we must stop sending packets. We add QUIC_FL_CONN_DRAINING connection flag to do so.	2022-04-06 15:52:35 +02:00
Frédéric Lécaille	eb2a2da67c	BUG/MINOR: quic: Missing TX packet deallocations Ensure all TX packets are deallocated. There may be remaining ones which will never be acknowledged or deemed lost.	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	b823bb7f7f	MINOR: quic: Add traces about list of frames This should be useful to have an idea of the list of frames which could be built towards the list of available frames when building packets. Same thing about the frames which could not be built because of a lack of room in the TX buffer.	2022-04-01 16:26:06 +02:00
Frédéric Lécaille	b002145e9f	MEDIUM: quic: Send ACK frames asap Due to a erroneous interpretation of the RFC 9000 (quic-transport), ACKs frames were always sent only after having received two ack-eliciting packets. This could trigger useless retransmissions for tail packets on the peer side. For now on, we send as soon as possible ACK frames as soon as we have ACK to send, in the same packets as the ack-eliciting frame packets, and we also send ACK frames after having received 2 ack-eliciting packets since the last time we sent an ACK frame with other ack-eliciting frames.	2022-04-01 16:22:52 +02:00
Frédéric Lécaille	205e4f359e	CLEANUP: quic: Remove all atomic operations on packet number spaces As such variables are handled by the QUIC connection I/O handler which runs always on the thread, there is no need to continue to use such atomic operations	2022-04-01 16:22:47 +02:00
Frédéric Lécaille	fc79006c92	CLEANUP: quic: Remove all atomic operations on quic_conn struct As the QUIC connections are always handled by the same thread there is no need anymore to continue to use atomic operations on such variables.	2022-04-01 16:22:44 +02:00
Amaury Denoyelle	d8e680cbaf	MEDIUM: mux-quic: remove qcs tree node The new qc_stream_desc type has a tree node for storage. Thus, we can remove the node in the qcs structure. When initializing a new stream, it is stored into the qcc streams_by_id tree. When the MUX releases it, it will freed as soon as its buffer is emptied. Before this, the quic-conn is responsible to store it inside its own streams_by_id tree.	2022-03-30 16:26:59 +02:00
Amaury Denoyelle	7272cd76fc	MEDIUM: quic: move transport fields from qcs to qc_conn_stream Move the xprt-buf and ack related fields from qcs to the qc_stream_desc structure. In exchange, qcs has a pointer to the low-level stream. For each new qcs, a qc_stream_desc is automatically allocated. This simplify the transport layer by removing qcs/mux manipulation during ACK frame parsing. An additional check is done to not notify the MUX on sending if the stream is already released : this case may now happen on retransmission. To complete this change, the quic_stream frame now references the quic_stream instance instead of a qcs.	2022-03-30 16:19:48 +02:00
Amaury Denoyelle	5c3859c509	MINOR: quic: implement stream descriptor for transport layer Currently, the mux qcs streams manage the Tx buffering, even after sending it to the transport layer. Buffers are emptied when acknowledgement are treated by the transport layer. This complicates the MUX liberation and we may loose some data after the MUX free. Change this paradigm by moving the buffering on the transport layer. For this goal, a new type is implemented as low-level stream at the transport layer, as a counterpart of qcs mux instances. This structure is called qc_stream_desc. This will allow to free the qcs/qcc instances without having to wait for acknowledge reception. For the moment, the quic-conn is responsible to store the qc_stream_desc in a new tree named streams_by_id. This will sligthly change in the next commits to remove the qcs node which has a similar purpose : qc_stream_desc instances will be shared between the qcc MUX and the quic-conn. This patch only introduces the new type definition and the function to manipulate it. The following commit will bring the rearchitecture in the qcs structure.	2022-03-30 16:16:07 +02:00
Amaury Denoyelle	cbc13b71c6	MINOR: mux-quic: define release app-ops Define a new callback release inside qcc_app_ops. It is called when the qcc MUX is freed via qc_release. This will allows to implement cleaning on the app layer.	2022-03-30 16:12:18 +02:00
Amaury Denoyelle	dccbd733f0	MINOR: mux-quic: reorganize qcs free Regroup some cleaning operations inside a new function qcs_free. This can be used for all streams, both through qcs_destroy and with uni-directional streams.	2022-03-30 16:12:18 +02:00
Amaury Denoyelle	50742294f5	MINOR: mux-quic: return qcs instance from qcc_get_qcs Refactoring on qcc_get_qcs : return the qcs instance instead of the tree node. This is useful to hide some eb64_entry macros for better readability.	2022-03-30 16:12:18 +02:00
William Lallemand	30fcca18a5	MINOR: ssl/lua: CertCache.set() allows to update an SSL certificate file The CertCache.set() function allows to update an SSL certificate file stored in the memory of the HAProxy process. This function does the same as "set ssl cert" + "commit ssl cert" over the CLI. This could be used to update the crt and key, as well as the OCSP, the SCTL, and the OSCP issuer. The implementation does yield every 10 ckch instances, the same way the "commit ssl cert" do.	2022-03-30 14:56:10 +02:00
William Lallemand	26654e7a59	MINOR: ssl: add "crt" in the cert_exts array The cert_exts array does handle "crt" the default way, however you might stil want to look for these extensions in the array.	2022-03-30 14:55:53 +02:00
William Lallemand	e60c7d6e59	MINOR: ssl: export ckch_inst_rebuild() ckch_inst_rebuild() will be needed to regenerate the ckch instances from the lua code, we need to export it.	2022-03-30 12:18:16 +02:00
William Lallemand	aaacc7e8ad	MINOR: ssl: move the cert_exts and the CERT_TYPE enum Move the cert_exts declaration and the CERT_TYPE enum in the .h in order to reuse them in another file.	2022-03-30 12:18:16 +02:00
William Lallemand	3b5a3a6c03	MINOR: ssl: split the cert commit io handler Extract the code that replace the ckch_store and its dependencies into the ckch_store_replace() function. This function must be used under the global ckch lock. It frees everything related to the old ckch_store.	2022-03-30 12:18:16 +02:00
Willy Tarreau	2100b383ab	MINOR: action: add a function to dump the list of actions for a ruleset The new function dump_act_rules() now dumps the list of actions supported by a ruleset. These actions are alphanumerically sorted first so that the produced output is easy to compare.	2022-03-30 11:19:22 +02:00
Willy Tarreau	3ff476e9ef	MINOR: tools: add strordered() to check whether strings are ordered When trying to sort sets of strings, it's often needed to required to compare 3 strings to see if the chosen one fits well between the two others. That's what this function does, in addition to being able to ignore extremities when they're NULL (typically for the first iteration for example).	2022-03-30 10:02:56 +02:00
Willy Tarreau	29d799d591	MINOR: sample: list registered sample converter functions Similar to the sample fetch keywords, let's also list the converter keywords. They're much simpler since there's no compatibility matrix. Instead the input and output types are listed. This is called by dump_registered_keywords() for the "cnv" keywords class.	2022-03-29 18:01:37 +02:00
Willy Tarreau	f78813f74f	MINOR: samples: add a function to list register sample fetch keywords New function smp_dump_fetch_kw lists registered sample fetch keywords with their compatibility matrix, mandatory and optional argument types, and output types. It's called from dump_registered_keywords() with class "smp".	2022-03-29 18:01:37 +02:00
Willy Tarreau	6ff7d1b9a5	MINOR: acl: add a function to dump the list of known ACL keywords New function acl_dump_kwd() dumps the registered ACL keywords and their sample-fetch equivalent to stdout. It's called by dump_registered_keywords() for keyword class "acl".	2022-03-29 18:01:37 +02:00
Willy Tarreau	06d0e2e034	MINOR: cli: add a new keyword dump function New function cli_list_keywords() scans the list of registered CLI keywords and dumps them on stdout. It's now called from dump_registered_keywords() for the class "cli". Some keywords are valid for the master, they'll be suffixed with "[MASTER]". Others are valid for the worker, they'll have "[WORKER]". Those accessible only in expert mode will show "[EXPERT]" and the experimental ones will show "[EXPERIM]".	2022-03-29 18:01:37 +02:00
Willy Tarreau	ca1acd6080	MINOR: config: add a function to dump all known config keywords All registered config keywords that are valid in the config parser are dumped to stdout organized like the regular sections (global, listen, etc). Some keywords that are known to only be valid in frontends or backends will be suffixed with [FE] or [BE]. All regularly registered "bind" and "server" keywords are also dumped, one per "bind" or "server" line. Those depending on ssl are listed after the "ssl" keyword. Doing so required to export the listener and server keyword lists that were static. The function is called from dump_registered_keywords() for keyword class "cfg".	2022-03-29 18:01:32 +02:00
Willy Tarreau	76871a4f8c	MINOR: management: add some basic keyword dump infrastructure It's difficult from outside haproxy to detect the supported keywords and syntax. Interestingly, many of our modern keywords are enumerated since they're registered from constructors, so it's not very hard to enumerate most of them. This patch creates some basic infrastructure to support dumping existing keywords from different classes on stdout. The format will differ depending on the classes, but the idea is that the output could easily be passed to a script that generates some simple syntax highlighting rules, completion rules for editors, syntax checkers or config parsers. The principle chosen here is that if "-dK" is passed on the command-line, at the end of the parsing the registered keywords will be dumped for the requested classes passed after "-dK". Special name "help" will show known classes, while "all" will execute all of them. The reason for doing that after the end of the config processor is that it will also enumerate internally-generated keywords, Lua or even those loaded from external code (e.g. if an add-on is loaded using LD_PRELOAD). A typical way to call this with a valid config would be: ./haproxy -dKall -q -c -f /path/to/config If there's no config available, feeding /dev/null will also do the job, though it will not be able to detect dynamically created keywords, of course. This patch also updates the management doc. For now nothing but the help is listed, various subsystems will follow in subsequent patches.	2022-03-29 17:55:54 +02:00
Amaury Denoyelle	0c2d964280	REORG: quic: use a dedicated quic_loss.c Move all inline functions with trace from quic_loss.h to a dedicated object file. This let to remove the TRACE_SOURCE macro definition outside of the include file. This change is required to be able to define another TRACE_SOUCE inside the mux_quic.c for a dedicated trace module.	2022-03-25 14:45:45 +01:00
Amaury Denoyelle	777969c163	BUILD: quic: add missing includes Complete the include list for the files quic_loss.h and quic_sock.c.	2022-03-25 14:45:45 +01:00
Amaury Denoyelle	1e5e5136ee	MINOR: mux-quic: support MAX_DATA frame parsing This commit is similar to the previous one but with MAX_DATA frames. This allows to increase the connection level flow-control limit. If the connection was blocked due to QC_CF_BLK_MFCTL flag, the flag is reseted.	2022-03-23 10:14:14 +01:00
Amaury Denoyelle	8727ff4668	MINOR: mux-quic: support MAX_STREAM_DATA frame parsing Implement a MUX method to parse MAX_STREAM_DATA. If the limit is greater than the previous one and the stream was blocked, the flag QC_SF_BLK_SFCTL is removed.	2022-03-23 10:09:39 +01:00
Amaury Denoyelle	05ce55e582	MEDIUM: mux-quic: respect peer connection data limit This commit is similar to the previous one, but this time on the connection level instead of the stream. When the connection limit is reached, the connection is flagged with QC_CF_BLK_MFCTL. This flag is checked in qc_send. qcs_push_frame uses a new parameter which is used to not exceed the connection flow-limit while calling it repeatdly over multiple streams instance before transfering data to the transport layer.	2022-03-23 10:05:29 +01:00
Amaury Denoyelle	6ea781919a	MEDIUM: mux-quic: respect peer bidirectional stream data limit Implement the flow-control max-streams-data limit on emission. We ensure that we never push more than the offset limit set by the peer. When the limit is reached, the stream is marked as blocked with a new flag QC_SF_BLK_SFCTL to disable emission. Currently, this is only implemented for bidirectional streams. It's required to unify the sending for unidirectional streams via qcs_push_frame from the H3 layer to respect the flow-control limit for them.	2022-03-23 10:05:29 +01:00
Amaury Denoyelle	78396e5ee8	MINOR: mux-quic: use shorter name for flow-control fields Rename the fields used for flow-control in the qcc structure. The objective is to have shorter name for better readability while keeping their purpose clear. It will be useful when the flow-control will be extended with new fields.	2022-03-23 10:05:29 +01:00
Amaury Denoyelle	1b4ebcb041	CLEANUP: mux-quic: adjust comment for coding-style Replace single-line comment style by /* ... */ format which is the standard for haproxy documentation.	2022-03-23 09:49:08 +01:00
Frédéric Lécaille	ce69cbc520	MINOR: quic: Add traces about stream TX buffer consumption This will be helpful to diagnose STREAM blocking states.	2022-03-23 09:01:45 +01:00
Dhruv Jain	1295798139	MEDIUM: mqtt: support mqtt_is_valid and mqtt_field_value converters for MQTTv3.1 In MQTTv3.1, protocol name is "MQIsdp" and protocol level is 3. The mqtt converters(mqtt_is_valid and mqtt_field_value) did not work for clients on mqttv3.1 because the mqtt_parse_connect() marked the CONNECT message invalid if either the protocol name is not "MQTT" or the protocol version is other than v3.1.1 or v5.0. To fix it, we have added the mqttv3.1 protocol name and version as part of the checks. This patch fixes the mqtt converters to support mqttv3.1 clients as well (issue #1600). It must be backported to 2.4.	2022-03-22 09:25:52 +01:00
Frédéric Lécaille	76fc07e9a0	BUG/MINOR: quic: Wrong TX packet related counters handling During the packet number space discarding, do no reset tx.in_flight counter before decrement it from other variables. Furthermore path prep_in_flight counter was not decremented.	2022-03-21 17:33:37 +01:00
Frédéric Lécaille	44ae75220a	BUG/MINOR: quic: Incorrect peer address validation We must consider the peer address as validated as soon as we received an handshake packet. An ACK frame in handshake packet was too restrictive. Rename the concerned flag to reflect this situation.	2022-03-21 14:27:09 +01:00
Frédéric Lécaille	2899fe2460	BUG/MINOR: quic: Missing TX packet initializations The most important one is the ->flags member which leads to an erratic xprt behavior. For instance a non ack-eliciting packet could be seen as ack-eliciting leading the xprt to try to retransmit a packet which are not ack-eliciting. In this case, the xprt does nothing and remains indefinitively in a blocking state.	2022-03-21 14:27:09 +01:00
Frédéric Lécaille	e2a1c1b372	MEDIUM: quic: Rework of the TX packets memory handling The TX packet refcounting had come with the multithreading support but not only. It is very useful to ease the management of the memory allocated for TX packets with TX frames attached to. At some locations of the code we have to move TX frames from a packet to a new one during retranmission when the packet has been deemed as lost or not. When deemed lost the memory allocated for the paquet must be released contrary to when its frames are retransmitted when probing (PTO). For now on, thanks to this patch we handle the TX packets memory this way. We increment the packet refcount when: - we insert it in its packet number space tree, - we attache an ack-eliciting frame to it. And reciprocally we decrement this refcount when: - we remove an ack-eliciting frame from the packet, - we delete the packet from its packet number space tree. Note that an optimization WOULD NOT be to fully reuse (without releasing its memorya TX packet to retransmit its contents (its ack-eliciting frames). Its information (timestamp, in flight length) to be processed by packet loss detection and the congestion control.	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	141982a4e1	MEDIUM: quic: Limit the number of ACK ranges When building a packet with an ACK frame, we store the largest acknowledged packet number sent in this frame in the packet (quic_tx_packet struc). When receiving an ack for such a packet we can purge the tree of acknowledged packet number ranges from the range sent before this largest acknowledged packet number.	2022-03-21 11:29:40 +01:00
Frédéric Lécaille	8f3ae0272f	CLEANUP: quic: "largest_acked_pn" pktns struc member moving This struct member stores the largest acked packet number which was received. It is used to build (TX) packet. But this is confusing to store it in the tx packet of the packet number space structure even if it is used to build and transmit packets.	2022-03-21 11:29:40 +01:00
Willy Tarreau	6a783e499c	MINOR: actions: add new function free_act_rule() to free a single rule There was free_act_rules() that frees all rules from a head but nothing to free a single rule. Currently some rulesets partially free their own rules on parsing error, and we're seeing some regtests emit errors under ASAN because of this. Let's first extract the code to free a rule into its own function so that it becomes possible to use it on a single rule.	2022-03-17 20:26:19 +01:00
Willy Tarreau	211ea252d9	BUG/MINOR: logs: fix logsrv leaks on clean exit Log servers are a real mess because: - entries are duplicated using memcpy() without their strings being reallocated, which results in these ones not being freeable every time. - a new field, ring_name, was added in 2.2 by commit `99c453df9` ("MEDIUM: ring: new section ring to declare custom ring buffers.") but it's never initialized during copies, causing the same issue - no attempt is made at freeing all that. Of course, running "haproxy -c" under ASAN quickly notices that and dumps a core. This patch adds the missing strdup() and initialization where required, adds a new free_logsrv() function to cleanly free() such a structure, calls it from the proxy when iterating over logsrvs instead of silently leaking their file names and ring names, and adds the same logsrv loop to the proxy_free_defaults() function so that we don't leak defaults sections on exit. It looks a bit entangled, but it comes as a whole because all this stuff is inter-dependent and was missing. It's probably preferable not to backport this in the foreseable future as it may reveal other jokes if some obscure parts continue to memcpy() the logsrv struct.	2022-03-17 19:53:46 +01:00
William Lallemand	0d05867e78	MINOR: server: export server_parse_sni_expr() function Export the server_parse_sni_expr() function in order to create a SNI expression in a server which was not parsed from the configuration.	2022-03-16 15:55:30 +01:00
William Lallemand	f5ba296ec8	CLEANUP: htx: remove unused co_htx_remove_blk() Remove the unused co_htx_remove_blk(), this function was confusing because you need to check the output size from the caller anyway.	2022-03-14 15:10:12 +01:00
Willy Tarreau	f1cb4ac745	BUG/MINOR: buffer: fix debugging condition in b_peek_varint() The BUG_ON_HOT() test condition added to b_peek_varint() by commit `8873b85bd` ("DEBUG: buf: add BUG_ON_HOT() to most buffer management functions") was wrong as <data> in this function is not b->data, so that was triggering during live dumps of H2 traces on the CLI when built with -DDEBUG_STRICT=2. No backport is needed.	2022-03-11 16:59:14 +01:00
Amaury Denoyelle	6ccfa3c40f	MEDIUM: mux-quic: improve bidir STREAM frames sending The current implementation of STREAM frames emission has some limitation. Most notably when we cannot sent all frames in a single qc_send run. In this case, frames are left in front of the MUX list. It will be re-send individually before other frames, possibly another frame from the same STREAM with new data. An opportunity to merge the frames is lost here. This method is now improved. If a frame cannot be send entirely, it is discarded. On the next qc_send run, we retry to send to this position. A new field qcs.sent_offset is used to remember this. A new frame list is used for each qc_send. The impact of this change is not precisely known. The most notable point is that it is a more logical method of emission. It might also improve performance as we do not keep old STREAM frames which might delay other streams.	2022-03-11 11:37:31 +01:00
Amaury Denoyelle	54445d04e4	MINOR: quic: implement sending confirmation Implement a new MUX function qcc_notify_send. This function must be called by the transport layer to confirm the sending of STREAM data to the MUX. For the moment, the function has no real purpose. However, it will be useful to solve limitations on push frame and implement the flow control.	2022-03-11 11:37:31 +01:00
Frédéric Lécaille	530601cd84	MEDIUM: quic: Implement the idle timeout feature The aim of the idle timeout is to silently closed the connection after a period of inactivity depending on the "max_idle_timeout" transport parameters advertised by the endpoints. We add a new task to implement this timer. Its expiry is updated each time we received an ack-eliciting packet, and each time we send an ack-eliciting packet if no other such packet was sent since we received the last ack-eliciting packet. Such conditions may be implemented thanks to QUIC_FL_CONN_IDLE_TIMER_RESTARTED_AFTER_READ new flag.	2022-03-11 11:37:30 +01:00
Frédéric Lécaille	c7a69e2aa5	MINOR: quic: Add a function to compute the current PTO There was not such a function at this time. This is needed to implement the idle timeout feature.	2022-03-11 11:37:30 +01:00
Frédéric Lécaille	12c169aaf0	BUG/MINOR: quic: ACK_REQUIRED and ACK_RECEIVED flag collision This packet number space flags were defined with the same value because defined at different places in the file. Assemble them at the same location with different values. This bug could unvalidate the peer address after it was validated during the handshake leading to the anti-amplication limit to be enabled again after having been disabled. The situation could not be unblocked (deadlock).	2022-03-11 11:37:30 +01:00
Frédéric Lécaille	f293b69521	MEDIUM: quic: Remove the QUIC connection reference counter There is no need to use such a reference counter anymore since the QUIC connections are always handled by the same thread. quic_conn_drop() is removed. Its code is merged into quic_conn_release().	2022-03-11 11:37:30 +01:00
Frédéric Lécaille	66d37fa051	MINOR: quic: Add max_idle_timeout advertisement handling When we store the remote transport parameters, we compute the maximum idle timeout for the connection which is the minimum of the two advertised max_idle_timeout transport parameter values if both have non-null values, or the maximum if one of the value is set and non-null.	2022-03-11 11:37:30 +01:00
Willy Tarreau	c6dae869ca	MINOR: rules: record the last http/tcp rule that gave a final verdict When a tcp-{request,response} content or http-request/http-response rule delivers a final verdict (deny, accept, redirect etc), the last evaluated one will now be recorded in the stream. The purpose is to permit to log the last one that performed a final action. For now the log is not produced.	2022-03-10 11:51:34 +01:00
Tim Duesterhus	b4b03779d0	MEDIUM: proxy: Store server_id_hdr_name as a `struct ist` The server_id_hdr_name is already processed as an ist in various locations lets also just store it as such. see `0643b0e7e` ("MINOR: proxy: Make `header_unique_id` a `struct ist`") for a very similar past commit.	2022-03-09 07:51:27 +01:00
Tim Duesterhus	e502c3e793	MINOR: proxy: Store orgto_hdr_name as a `struct ist` The orgto_hdr_name is already processed as an ist in `http_process_request`, lets also just store it as such. see `0643b0e7e` ("MINOR: proxy: Make `header_unique_id` a `struct ist`") for a very similar past commit.	2022-03-09 07:51:27 +01:00
Tim Duesterhus	b50ab8489e	MINOR: proxy: Store fwdfor_hdr_name as a `struct ist` The fwdfor_hdr_name is already processed as an ist in `http_process_request`, lets also just store it as such. see `0643b0e7e` ("MINOR: proxy: Make `header_unique_id` a `struct ist`") for a very similar past commit.	2022-03-09 07:51:27 +01:00
Tim Duesterhus	4b1fcaaee3	MINOR: proxy: Store monitor_uri as a `struct ist` The monitor_uri is already processed as an ist in `http_wait_for_request`, lets also just store it as such. see `0643b0e7e` ("MINOR: proxy: Make `header_unique_id` a `struct ist`") for a very similar past commit.	2022-03-09 07:51:27 +01:00
David Carlier	6709538068	BUILD: fix recent build breakage of freebsd caused by kFreeBSD build fix Supporting kFreebsd previously led to FreeBSD (< 14) build breakage: In file included from src/cpuset.c:5: In file included from include/haproxy/cpuset.h:4: include/haproxy/cpuset-t.h:46:2: error: unknown type name 'cpu_set_t'; did you mean 'cpuset_t'? CPUSET_REPR cpuset; ^~~~~~~~~~~ cpuset_t include/haproxy/cpuset-t.h:21:22: note: expanded from macro 'CPUSET_REPR' # define CPUSET_REPR cpu_set_t ^	2022-03-08 16:03:28 +01:00
Frédéric Lécaille	5bcfd33063	BUG/MAJOR: quic: Wrong quic_max_available_room() returned value Around limits for QUIC integer encoding, this functions could return wrong values which lead to qc_build_frms() to prepare wrong CRYPTO (less chances) or STREAM frames (more chances). qc_do_build_pkt() could build wrong packets with bad CRYPTO/STREAM frames which could not be decoded by the peer. In such a case ngtcp2 closes the connection with an ENCRYPTION_ERROR error in a transport CONNECTION_CLOSE frame.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	4fe7d8a5b2	MINOR: quic: Add quic_max_int_by_size() function This function returns the maximum integer which may be encoded with a number of bytes passed as parameter. Useful to precisely compute the number of bytes which may used to fulfill a buffer with lengths as QUIC enteger encoded prefixes for the number of following bytes.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	9777ead2ed	CLEANUP: quic: Remove window redundant variable from NewReno algorithm state struct We use the window variable which is stored in the path struct.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	0e7c9a7143	MINOR: quic: More precise window update calculation When in congestion avoidance state and when acknowledging an <acked> number bytes we must increase the congestion window by at most one datagram (<path->mtu>) by congestion window. So thanks to this patch we apply a ratio to the current number of acked bytes : <acked> * <path->mtu> / <cwnd>. So, when <cwnd> bytes are acked we precisely increment <cwnd> by <path->mtu>. Furthermore we take into an account the number of remaining acknowledged bytes each time we increment the window by <acked> storing their values in the algorithm struct state (->remain_acked) so that it might be take into an account at the next ACK event.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	abdf4a1533	BUG/MINOR: quic: Confusion betwen "in_flight" and "prep_in_flight" in quic_path_prep_data() This function returns the remaining number of bytes which can be sent on the network before fulfilling the congestion window. There is a counter for the number of prepared data and another one for the really in flight number of bytes (in_flight). These variable have been mixed up.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	5f6783094d	CLEANUP: quic: Remove useless definitions from quic_cc_event struct Since the persistent congestion detection is done out of the congestion controllers, there is no need to pass them information through quic_cc_event struct. We remove its useless members. Also remove qc_cc_loss_event() which is no more used.	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	a5ee0ae6a2	MINOR: quic: Persistent congestion detection outside of controllers We establish the persistent congestion out of any congestion controller to improve the algorithms genericity. This path characteristic detection may be implemented regarless of the underlying congestion control algorithm. Send congestion (loss) event using directly quic_cc_event(), so without qc_cc_loss_event() wrapper function around quic_cc_event(). Take the opportunity of this patch to shorten "newest_time_sent" member field of quic_cc_event to "time_sent".	2022-03-04 17:47:32 +01:00
Frédéric Lécaille	83bfca6c71	MINOR: quic: Add a "slow start" callback to congestion controller We want to be able to make the congestion controllers re-enter the slow start state outside of the congestion controllers themselves. So, we add a callback ->slow_start() to do so. Define this callback for NewReno algorithm.	2022-03-04 17:47:32 +01:00
David Carlier	43a568575f	BUILD: fix kFreeBSD build. kFreeBSD needs to be treated as a distinct target from FreeBSD since the underlying system libc is the GNU one. Thus, relying only on __GLIBC__ no longer suffice. - freebsd-glibc new target, key difference is including crypt.h and linking to libdl like linux. - cpu affinity available but the api is still the FreeBSD's. - enabling auxiliary data access only for Linux. Patch based on preliminary work done by @bigon. closes #1555	2022-03-04 17:19:12 +01:00
Amaury Denoyelle	c055e30176	MEDIUM: mux-quic: implement MAX_STREAMS emission for bidir streams Implement the locally flow-control streams limit for opened bidirectional streams. Add a counter which is used to count the total number of closed streams. If this number is big enough, emit a MAX_STREAMS frame to increase the limit of remotely opened bidirectional streams. This is the first commit to implement QUIC flow-control. A series of patches should follow to complete this. This is required to be able to handle more than 100 client requests. This should help to validate the Multiplexing interop test.	2022-03-04 17:00:12 +01:00
Amaury Denoyelle	2c71fe58f0	MEDIUM: mux-quic: use direct send transport API for STREAMs Modify the STREAM emission in qc_send. Use the new transport function qc_send_app_pkts to directly send the list of constructed frames. This allows to remove the tasklet wakeup on the quic_conn and should reduce the latency. If not all frames are send after the transport call, subscribe the MUX on the lower layer to be able to retry. Currently there is a bug because the transport layer does not retry to send frames in excess after a successful sendto. This might cause the transfer to be interrupted.	2022-03-04 17:00:12 +01:00
Amaury Denoyelle	414df7684a	MINOR: mux-quic: define new unions for flow-control fields Define two new unions in the qcc structure named 'lfctl' and 'rfctl'. For the moment they are empty. They will be completed to store the initial and current level for flow-control on the local and remote side.	2022-03-04 17:00:12 +01:00
Amaury Denoyelle	0dc40f06d1	MINOR: mux-quic: complete functions to detect stream type Improve the functions used to detect the stream characteristics : uni/bidirectional and local/remote initiated. Most notably, these functions are now designed to work transparently for a MUX in the frontend or backend side. For this, we use the connection to determine the current MUX side. This will be useful if QUIC is implemented on the server side.	2022-03-04 17:00:12 +01:00
Amaury Denoyelle	749cb647b1	MINOR: mux-quic: refactor transport parameters init Since QUIC accept handling has been improved, the MUX is initialized after the handshake completion. Thus its safe to access transport parameters in qc_init via the quic_conn. Remove quic_mux_transport_params_update which was called by the transport for the MUX. This improves the architecture by removing a direct call from the transport to the MUX. The deleted function body is not transfered to qc_init because this part will change heavily in the near future when implementing the flow-control.	2022-03-04 17:00:12 +01:00
Frédéric Lécaille	c2f561ce1e	MINOR: quic: Export qc_send_app_pkts() This is at least to make this function be callable by the mux.	2022-03-04 17:00:12 +01:00
Willy Tarreau	3dfb7da04b	CLEANUP: tree-wide: remove a few rare non-ASCII chars As reported by Tim in issue #1428, our sources are clean, there are just a few files with a few rare non-ASCII chars for the paragraph symbol, a few typos, or in Fred's name. Given that Fred already uses the non-accentuated form at other places like on the public list, let's uniformize all this and make sure the code displays equally everywhere.	2022-03-04 08:58:32 +01:00
Willy Tarreau	e81248c0c8	BUG/MINOR: pool: always align pool_heads to 64 bytes This is the pool equivalent of commit `97ea9c49f` ("BUG/MEDIUM: fd: always align fdtab[] to 64 bytes"). After a careful code review, it happens that the pool heads are the other structures allocated with malloc/calloc that claim to be aligned to a size larger than what the allocator can offer. While no issue was reported on them, no memset() is performed and no type is large, this is a problem waiting to happen, so better fix it. In addition, it's relatively easy to do by storing the allocation address inside the pool_head itself and use it at free() time. Finally, threads might benefit from the fact that the caches will really be aligned and that there will be no false sharing. This should be backported to all versions where it applies easily.	2022-03-02 18:22:08 +01:00
Willy Tarreau	06e66c84fc	DEBUG: reduce the footprint of BUG_ON() calls Many inline functions involve some BUG_ON() calls and because of the partial complexity of the functions, they're not inlined anymore (e.g. co_data()). The reason is that the expression instantiates the message, its size, sometimes a counter, then the atomic OR to taint the process, and the back trace. That can be a lot for an inline function and most of it is always the same. This commit modifies this by delegating the common parts to a dedicated function "complain()" that takes care of updating the counter if needed, writing the message and measuring its length, and tainting the process. This way the caller only has to check a condition, pass a pointer to the preset message, and the info about the type (bug or warn) for the tainting, then decide whether to dump or crash. Note that this part could also be moved to the function but resulted in complain() always being at the top of the stack, which didn't seem like an improvement. Thanks to these changes, the BUG_ON() calls do not result in uninlining functions anymore and the overall code size was reduced by 60 to 120 kB depending on the build options.	2022-03-02 16:00:42 +01:00
Willy Tarreau	a631b86523	BUILD: tcpcheck: do not declare tcp_check_keywords_register() inline This one is referenced in initcalls by its pointer, it makes no sense to declare it inline. At best it causes function duplication, at worst it doesn't build on older compilers.	2022-03-02 14:54:44 +01:00
Willy Tarreau	4de2cda104	BUILD: trace: do not declare trace_registre_source() inline This one is referenced in initcalls by its pointer, it makes no sense to declare it inline. At best it causes function duplication, at worst it doesn't build on older compilers.	2022-03-02 14:53:00 +01:00
Willy Tarreau	368479c3fc	BUILD: http_rules: do not declare http_*_keywords_registre() inline The 3 functions http_{req,res,after_res}_keywords_register() are referenced in initcalls by their pointer, it makes no sense to declare them inline. At best it causes function duplication, at worst it doesn't build on older compilers.	2022-03-02 14:50:38 +01:00
Willy Tarreau	d318e4e022	BUILD: connection: do not declare register_mux_proto() inline This one is referenced in initcalls by its pointer, it makes no sense to declare it inline. At best it causes function duplication, at worst it doesn't build on older compilers.	2022-03-02 14:46:45 +01:00
Willy Tarreau	e4149cdbc6	BUILD: conn_stream: avoid null-deref warnings on gcc 6 gcc 6 continues its saga with excessive reports of null-deref warnings. This time it was in the IS_HTX_CS() macro. Let's use __cs_conn() after cs_conn() was checked.	2022-03-02 14:39:39 +01:00
Frédéric Lécaille	bd24208673	MINOR: quic: Assemble QUIC TLS flags at the same level Do not distinguish the direction (TX/RX) when settings TLS secrets flags. There is not such a distinction in the RFC 9001. Assemble them at the same level: at the upper context level.	2022-03-01 16:34:03 +01:00
Frédéric Lécaille	00e2400fa6	MINOR: quic: Post handshake I/O callback switching Implement a simple task quic_conn_app_io_cb() to be used after the handshakes have completed.	2022-03-01 16:22:35 +01:00
Frédéric Lécaille	5757b4a50e	MINOR: quic: Ensure PTO timer is not set in the past Wakeup asap the timer task when setting its timer in the past. Take also the opportunity of this patch to make simplify quic_pto_pktns(): calling tick_first() is useless here to compare <lpto> with <tmp_pto>.	2022-03-01 16:22:35 +01:00
Julien Thomas	59e6bcdcea	BUILD: ssl: another build warning on LIBRESSL_VERSION_NUMBER We had several warnings when building haproxy 2.5.4 with old openssl 1.0.1e. This version of openssl is the latest available in EOL centos 6. include/haproxy/openssl-compat.h:157:51: \ warning: "LIBRESSL_VERSION_NUMBER" is not defined This patch fixed the build. It changes the #if condition, as done in other similar parts of openssl-compat.h.	2022-03-01 15:49:30 +01:00
Amaury Denoyelle	0e3010b1bb	MEDIUM: quic: rearchitecture Rx path for bidirectional STREAM frames Reorganize the Rx path for STREAM frames on bidirectional streams. A new function qcc_recv is implemented on the MUX. It will handle the STREAM frames copy and offset calculation from transport to MUX. Another function named qcc_decode_qcs from the MUX can be called by transport each time new STREAM data has been copied. The architecture is now cleaner with the MUX layer in charge of parsing the STREAM frames offsets. This is required to be able to implement the flow-control on the MUX layer. Note that as a convenience, a STREAM frame is not partially copied to the MUX buffer. This simplify the implementation for the moment but it may change in the future to optimize the STREAM frames handling. For the moment, only bidirectional streams benefit from this change. In the future, it may be extended to unidirectional streams to unify the STREAM frames processing.	2022-03-01 11:07:27 +01:00
Amaury Denoyelle	3c4303998f	BUG/MINOR: quic: support FIN on Rx-buffered STREAM frames FIN flag on a STREAM frame was not detected if the frame was previously buffered on qcs.rx.frms before being handled. To fix this, copy the fin field from the quic_stream instance to quic_rx_strm_frm. This is required to properly notify the FIN flag on qc_treat_rx_strm_frms for the MUX layer. Without this fix, the request channel might be left opened after the last STREAM frame reception if there is out-of-order frames on the Rx path.	2022-03-01 11:07:06 +01:00
Amaury Denoyelle	3bf06093dc	MINOR: mux-quic: define flag for last received frame This flag is set when the STREAM frame with FIN set has been received on a qcs instance. For now, this is only used as a BUG_ON guard to prevent against multiple frames with FIN set. It will also be useful when reorganize the RX path and move some of its code in the mux.	2022-03-01 10:52:31 +01:00
Willy Tarreau	5a001a0e4d	BUILD: debug: fix build warning on older compilers around DEBUG_STRICT_ACTION The new macro was introduced with commit `86bcc5308` ("DEBUG: implement 4 levels of choices between warn and crash.") but some older compilers can complain that we test the value when the macro is not defined despite having already been checked in a previous #if directive. Let's just repeat the test for the definition.	2022-02-28 17:59:28 +01:00
Christopher Faulet	8bc1759f60	DEBUG: stream-int: Fix BUG_ON used to test appctx in si_applet_ops callbacks `693b23bb1` ("MEDIUM: tree-wide: Use unsafe conn-stream API when it is relevant") introduced a regression in DEBUG_STRICT mode because some BUG_ON conditions were inverted. It should ok now. In addition, ALREADY_CHECKED macro was removed from appctx_wakeup() function because it is useless now.	2022-02-28 17:29:11 +01:00
Christopher Faulet	9936dc6577	REORG: stream-int: Uninline si_sync_recv() and make si_cs_recv() private This way si__recv() and si__sned() API are defined the same way. si_sync_snd/si_sync_recv are both exported and defined in the C file. And si_cs_send/si_cs_recv are private and only used by stream-interface internals.	2022-02-28 17:16:47 +01:00
Christopher Faulet	693b23bb10	MEDIUM: tree-wide: Use unsafe conn-stream API when it is relevant The unsafe conn-stream API (__cs_*) is now used when we are sure the good endpoint or application is attached to the conn-stream. This avoids compiler warnings about possible null derefs. It also simplify the code and clear up any ambiguity about manipulated entities.	2022-02-28 17:13:36 +01:00
Christopher Faulet	e645d88c6b	MINOR: conn-stream: Improve API to have safe/unsafe accessors Depending on the context, we know the endpoint or the application attached to the conn_stream is defined and we know its type. However, having accessors testing the endpoint or the application may lead the compiler to report possible null derefs here and there. The alternative is to add useless tests or use ALREAD_CHECKED/DISGUISE macros. It is tedious and inelegant. So now, similarily to the ob API, the safe API, testing endpoint/application, relies on an unsafe one (same name prefixed with '__'). This way, any caller may use the unsafe API when it is relevant. In addition, there is no reason to test the conn-stream itself. It is the caller responsibility to be sure there is a conn-stream to get its endpoint or its application. And most of type, we are sure to have a conn-stream.	2022-02-28 17:13:36 +01:00
Willy Tarreau	68ae291cd2	DEBUG: channel: add consistency checks using BUG_ON_HOT() in some key functions A few functions such as c_adv(), c_rew(), co_set_data() or co_skip() got a BUG_ON_HOT() to make sure they're not used to push more data than available in the buffer. Note that with HTX the margin can be high and will less easily trigger, but the goal is to detect a misuse early enough. co_data() should never be called with a wrong c->output. At least it never happens in regtests, but we're adding a CHECK_IF_HOT() there to avoid crashing but report it if it ever happened when the hot path checks are enabled.	2022-02-28 16:59:17 +01:00
Willy Tarreau	84240044f0	MINOR: channel: don't use co_set_data() to decrement output The use of co_set_data() should be strictly limited to setting the amount of existing data to be transmitted. It ought not be used to decrement the output after the data have left the buffer, because doing so involves performing incorrect calculations using co_data() that still comprises data that are not in the buffer anymore. Let's use c_rew() for this, which is made exactly for this purpose, i.e. decrement c->output by as much as requested. This is cleaner, faster, and will permit stricter checks.	2022-02-28 16:51:23 +01:00
Willy Tarreau	8873b85bd9	DEBUG: buf: add BUG_ON_HOT() to most buffer management functions A number of tests are now performed in low-level buffer management functions to verify that we're not appending data to a full buffer for example, or that the buffer passed in argument is consistent in that its data don't outweigh its size. The few functions that already involve memcpy() or memmove() instead got a BUG_ON() that will always be enabled, since the overhead remains minimalist.	2022-02-28 16:14:02 +01:00
Willy Tarreau	a8f4b34bb7	DEBUG: buf: replace some sensitive BUG_ON() with BUG_ON_HOT() The buffer ring management functions br_* were all stuffed with BUG_ON() statements that never triggered and that are on some fast paths (e.g. in mux_h2). Let's turn them to BUG_ON_HOT() instead.	2022-02-28 16:10:00 +01:00
Willy Tarreau	7bd7954535	DEBUG: add two new macros to enable debugging in hot paths Two new BUG_ON variants, BUG_ON_HOT() and CHECK_IF_HOT() are introduced to debug hot paths (such as low-level API functions). These ones must not be enabled by default as they would significantly affect performance but they may be enabled by setting DEBUG_STRICT to a value above 1. In this case, DEBUG_STRICT_ACTION is mostly respected with a small change, which is that the no_crash variant of BUG_ON() isn't turned to a regular warning but to a one-time warning so as not to spam with warnings in a hot path. It is for this reason that there is no WARN_ON_HOT().	2022-02-28 15:32:24 +01:00
Willy Tarreau	86bcc53084	DEBUG: implement 4 levels of choices between warn and crash. We used to have DEBUG_STRICT_NOCRASH to disable crashes on BUG_ON(). Now we have other levels (WARN_ON(), CHECK_IF()) so we need something finer-grained. This patch introduces DEBUG_STRICT_ACTION which takes an integer value. 0 disables crashes and is the equivalent of DEBUG_STRICT_NOCRASH. 1 is the default and only enables crashes on BUG_ON(). 2 also enables crashes on WARN_ON(), and 3 also enables warnings on CHECK_IF(), and is suited to developers and CI.	2022-02-28 15:00:55 +01:00
Willy Tarreau	ef16578822	DEBUG: improve BUG_ON output message accuracy Now we'll explicitly mention if the test was a bug/warn/check, and "FATAL" is only displayed when the process crashes. The non-crashing BUG_ON() also suggests to report to developers.	2022-02-28 15:00:03 +01:00
Willy Tarreau	6d3f1e322e	DEBUG: rename WARN_ON_ONCE() to CHECK_IF() The only reason for warning once is to check if a condition really happens. Let's use a term that better translates the intent, that's important when reading the code.	2022-02-28 11:51:23 +01:00
Amaury Denoyelle	642ab06313	MINOR: quic: adjust buffer handling for STREAM transmission Simplify the data manipulation of STREAM frames on TX. Only stream data and len field are used to generate a valid STREAM frames from the buffer. Do not use the offset field, which required that a single buffer instance should be shared for every frames on a single stream.	2022-02-25 15:06:17 +01:00
Willy Tarreau	897c861aea	DEBUG: report BUG_ON() and WARN_ON() in the tainted flags It can be useful to know from the "tainted" variable whether any WARN_ON() or BUG_ON() triggered. Both were now added.	2022-02-25 11:55:47 +01:00
Willy Tarreau	4e0a8b1224	DEBUG: add a new WARN_ON_ONCE() macro This one will maintain a static counter per call place and will only emit the warning on the first call. It may be used to invite users to report an unexpected event without spamming them with messages.	2022-02-25 11:55:47 +01:00
Willy Tarreau	a79db30c63	DEBUG: make the _BUG_ON() macro return the condition By doing so it now becomes an expression and will allow for example to use WARN_ON() in tests, for example: if (WARN_ON(cond)) return NULL;	2022-02-25 11:55:47 +01:00
Willy Tarreau	305cfbde43	DBEUG: add a new WARN_ON() macro This is the same as BUG_ON() except that it never crashes and only emits a warning and a backtrace, inviting users to report the problem. This will be usable for non-fatal issues that should not happen and need to be fixed. This way the BUG_ON() when using DEBUG_STRICT_NOCRASH is effectively an equivalent of WARN_ON().	2022-02-25 11:55:47 +01:00
Willy Tarreau	f19aab88d5	DEBUG: mark ABORT_NOW() as unreachable The purpose is to make the program die at this point, so let's help the compiler optimise the code (especially in sensitive areas) by telling it that ABORT_NOW() does not return. This reduces the overall code size by ~0.5%.	2022-02-25 11:55:47 +01:00
Willy Tarreau	be0dbba6ec	DEBUG: cleanup BUG_ON() configuration The BUG_ON() macro handling is complicated because it relies on a conditional CRASH_NOW() macro whose definition depends on DEBUG_STRICT and DEBUG_STRICT_NOCRASH. Let's rethink the whole thing differently, and instead make the underlying _BUG_ON() macro take a crash argument to decide whether to crash or not, as well as a prefix and a suffix for the message, that will allow to distinguish between variants. Now the suffix is set to a message explaining we don't crash when needed. This also allows to get rid of the CRASH_NOW() macro and to define much simpler new macros.	2022-02-25 11:55:47 +01:00
Willy Tarreau	1ea8bc4c48	DEBUG: cleanup back trace generation Most BUG()/ABORT() macros were duplicating the same code to call the backtrace production to stderr, better place that into a new DUMP_TRACE() macro.	2022-02-25 11:55:47 +01:00
Willy Tarreau	edd426871f	DEBUG: move the tainted stuff to bug.h for easier inclusion The functions needed to manipulate the "tainted" flags were located in too high a level to be callable from the lower code layers. Let's move them to bug.h.	2022-02-25 11:55:38 +01:00
Christopher Faulet	2da02ae8b2	BUILD: tree-wide: Avoid warnings about undefined entities retrieved from a CS Since recent changes related to the conn-stream/stream-interface refactoring, GCC reports potential null pointer dereferences when we get the appctx, the stream or the stream-interface from the conn-strem. Of course, depending on the time, these entities may be null. But at many places, we know they are defined and it is safe to get them without any check. Thus, we use ALREADY_CHECKED() macro to silent these warnings. Note that the refactoring is unfinished, so it is not a real issue for now.	2022-02-24 13:56:52 +01:00
Christopher Faulet	c983b2114d	CLEANUP: backend: Don't export connect_server anymore connect_server() function is only called from backend.c. So make it static.	2022-02-24 11:00:03 +01:00
Christopher Faulet	e3a3af1ec8	CLEANUP: conn-stream: Remove cs_destroy() This function is no longer used.	2022-02-24 11:00:03 +01:00
Christopher Faulet	c36de9dc93	MINOR: conn-stream: Release a CS when both app and endp are detached cs_detach_app() function is added to detach an app from a conn-stream. And now, both cs_detach_app() and cs_detach_endp() release the conn-stream when both the app and the endpoint are detached.	2022-02-24 11:00:03 +01:00
Christopher Faulet	014ac35eb2	CLEANUP: stream-int: rename si_reset() to si_init() si_reset() function is only used when a stream-interface is allocated. Thus rename it to si_init() insteaad.	2022-02-24 11:00:03 +01:00
Christopher Faulet	cda94accb1	MAJOR: stream/conn_stream: Move the stream-interface into the conn-stream Thanks to all previous changes, it is now possible to move the stream-interface into the conn-stream. To do so, some SI functions are removed and their conn-stream counterparts are added. In addition, the conn-stream is now responsible to create and release the stream-interface. While the stream-interfaces were inlined in the stream structure, there is now a pointer in the conn-stream. stream-interfaces are now dynamically allocated. Thus a dedicated pool is added. It is a temporary change because, at the end, the stream-interface structure will most probably disappear.	2022-02-24 11:00:03 +01:00
Christopher Faulet	9a86f6399f	CLEANUP: conn-stream: Don't export conn-stream pool There is no reason to export the conn-stream pool.	2022-02-24 11:00:03 +01:00
Christopher Faulet	a73c9f0faa	MINOR: conn-stream: Rename cs_detach() to cs_detach_endp() Because cs_detach() is releated to the endpoint only, the function is renamed. The main purpose of this patch is to be able to add a function to detach the conn-stream from the application.	2022-02-24 11:00:02 +01:00
Christopher Faulet	5c8b47f665	MINOR: stream: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the stream part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	165ca0e812	MINOR: stream-int: Always access the stream-int via the conn-stream To be able to move the stream-interface from the stream to the conn-stream, all access to the SI is done via the conn-stream. This patch is limited to the stream-interface part.	2022-02-24 11:00:02 +01:00
Christopher Faulet	95a61e8a0e	MINOR: stream: Add pointer to front/back conn-streams into stream struct frontend and backend conn-streams are now directly accesible from the stream. This way, and with some other changes, it will be possible to remove the stream-interfaces from the stream structure.	2022-02-24 11:00:02 +01:00
Christopher Faulet	f835dea939	MEDIUM: conn_stream: Add a pointer to the app object into the conn-stream In the same way the conn-stream has a pointer to the stream endpoint , this patch adds a pointer to the application entity in the conn-stream structure. For now, it is a stream or a health-check. It is mandatory to merge the stream-interface with the conn-stream.	2022-02-24 11:00:02 +01:00
Christopher Faulet	86e1c3381b	MEDIUM: applet: Set the conn-stream as appctx owner instead of the stream-int Because appctx is now an endpoint of the conn-stream, there is no reason to still have the stream-interface as appctx owner. Thus, the conn-stream is now the appctx owner.	2022-02-24 11:00:02 +01:00
Christopher Faulet	13a35e5752	MAJOR: conn_stream/stream-int: move the appctx to the conn-stream Thanks to previous changes, it is now possible to set an appctx as endpoint for a conn-stream. This means the appctx is no longer linked to the stream-interface but to the conn-stream. Thus, a pointer to the conn-stream is explicitly stored in the stream-interface. The endpoint (connection or appctx) can be retrieved via the conn-stream.	2022-02-24 11:00:02 +01:00
Christopher Faulet	dd2d0d8b80	MEDIUM: conn-stream: Be prepared to use an appctx as conn-stream endpoint To be able to use an appctx as conn-stream endpoint, the connection is no longer stored as is in the conn-stream. The obj-type is used instead.	2022-02-24 11:00:02 +01:00
Christopher Faulet	897d612d68	MEDIUM: conn-stream: No longer access connection field directly To be able to handle applets as a conn-stream endpoint, we must be prepared to handle different types of endpoints. First of all, the conn-strream's connection must no longer be used directly.	2022-02-24 11:00:02 +01:00
Christopher Faulet	1329f2a12a	REORG: conn_stream: move conn-stream stuff in dedicated files Move code dealing with the conn-streams in dedicated files.	2022-02-24 11:00:02 +01:00
Christopher Faulet	e00ad358c9	MEDIUM: stream: No longer release backend conn-stream on connection retry The backend conn-stream is no longer released on connection retry. This means the conn-stream is detached from the underlying connection but not released. Thus, during connection retries, the stream has always an allocated conn-stream with no connection. All previous changes were made to make this possible. Note that .attach() mux callback function was changed to get the conn-stream as argument. The muxes are no longer responsible to create the conn-stream when a server connection is attached to a stream.	2022-02-24 11:00:02 +01:00
Christopher Faulet	e39827de0d	MINOR: stream-int: Be able to allocate a CS without connection si_alloc_cs() function may now be called without connection. It is mandatory to allocate the backend conn-stream during the stream creation.	2022-02-24 11:00:02 +01:00
Christopher Faulet	1a3b598b47	MINOR: stream-int: Add function to attach a connection to a SI si_attach_conn() function should be used to attach a connection to a stream-interface. It created a conn-stream if necessary. This function is mandatory to be able to keep the backend conn-stream during connection retries.	2022-02-24 11:00:02 +01:00
Christopher Faulet	20a6501051	MINOR: stream-int: Add function to reset a SI endpoint si_reset_endpoint() function may be used to reset the SI's endpoint without releasing the conn-stream if the endpoint is a connection. If the endpoint is an appctx, it is released. This change is mandatory to merge the SI and the CS and keep the backend conn-stream attached to the stream during connection retries.	2022-02-24 11:00:02 +01:00
Christopher Faulet	2b4e8b7b2d	MINOR: connection: Add a function to detach a conn-stream from the connection cs_detach() function is added to detach a conn-stream from the underlying connection. This part will evovle to handle applets too. Concretely, cs_destroy() is split to detach the conn-stream from its endpoint, via cs_detach(), and then, the conn-stream is released, via cs_free().	2022-02-24 11:00:01 +01:00
Christopher Faulet	0256da14a5	MINOR: connection: Be prepared to handle conn-stream with no connection The conn-stream will progressively replace the stream-interface. Thus, a stream will have to allocate the backend conn-stream during its creation. This means it will be possible to have a conn-stream with no connection. To prepare this change, we test the conn-stream's connection when we retrieve it.	2022-02-24 11:00:01 +01:00
Christopher Faulet	719ceef79c	MINOR: stream-int: Handle appctx case first when releasing the endpoint Stream-interfaces will be moved in the conn-stream and the appctx will be moved at the same level than the muxes. Idea is to merge the stream-interface and the conn-stream and have a better symmetry between the muxes and the applets. To limit bugs during this refactoring, when the SI endpoint is released, the appctx case is handled first.	2022-02-24 11:00:01 +01:00
Willy Tarreau	1408b1f8be	MINOR: pools: delegate parsing of command line option -dM to a new function New function pool_parse_debugging() is now dedicated to parsing options of -dM. For now it only handles the optional memory poisonning byte, but the function may already return an informative message to be printed for help, a warning or an error. This way we'll reuse it for the settings that will be needed for configurable debugging options.	2022-02-23 17:28:41 +01:00
Willy Tarreau	3ebe4d989c	MEDIUM: initcall: move STG_REGISTER earlier The STG_REGISTER init level is used to register known keywords and protocol stacks. It must be called earlier because some of the init code already relies on it to be known. For example, "haproxy -vv" for now is constrained to start very late only because of this. This patch moves it between STG_LOCK and STG_ALLOC, which is fine as it's used for static registration.	2022-02-23 17:11:33 +01:00
Willy Tarreau	ef301b7556	MINOR: pools: add a debugging flag for memory poisonning option Now -dM will set POOL_DBG_POISON for consistency with the rest of the pool debugging options. As such now we only check for the new flag, which allows the default value to be preset.	2022-02-23 17:11:33 +01:00
Willy Tarreau	13d7775b06	MINOR: pools: replace DEBUG_MEMORY_POOLS with runtime POOL_DBG_TAG This option used to allow to store a marker at the end of the area, which was used as a canary and detection against wrong freeing while the object is used, and as a pointer to the last pool_free() caller when back in cache. Now that we can compute the offsets at runtime, let's check it at run time and continue the code simplification.	2022-02-23 17:11:33 +01:00
Willy Tarreau	0271822f17	MINOR: pools: replace DEBUG_POOL_TRACING with runtime POOL_DBG_CALLER This option used to allow to store a pointer to the caller of the last pool_alloc() or pool_free() at the end of the area. Now that we can compute the offsets at runtime, let's check it at run time and continue the code simplification. In __pool_alloc() we now always calculate the return address (which is quite cheap), and the POOL_DEBUG_TRACE_CALLER() calls are conditionned on the status of debugging option.	2022-02-23 17:11:33 +01:00
Willy Tarreau	42705d06b7	MINOR: pools: get rid of POOL_EXTRA This macro is build-time dependent and is almost unused, yet where it cannot easily be avoided. Now that we store the distinction between pool->size and pool->alloc_sz, we don't need to maintain it and we can instead compute it on the fly when creating a pool. This is what this patch does. The variables are for now pretty static, but this is sufficient to kill the macro and will allow to set them more dynamically.	2022-02-23 17:11:33 +01:00
Willy Tarreau	96d5bc7379	MINOR: pools: store the allocated size for each pool The allocated size is the visible size plus the extra storage. Since for now we can store up to two extra elements (mark and tracer), it's convenient because now we know that the mark is always stored at ->size, and the tracer is always before ->alloc_sz.	2022-02-23 17:11:33 +01:00
Willy Tarreau	e981631d27	MEDIUM: pools: replace CONFIG_HAP_POOLS with a runtime "NO_CACHE" flag. Like previous patches, this replaces the build-time code paths that were conditionned by CONFIG_HAP_POOLS with runtime paths conditionned by !POOL_DBG_NO_CACHE. One trivial test had to be added in the hot path in __pool_alloc() to refrain from calling pool_get_from_cache(), and another one in __pool_free() to avoid calling pool_put_to_cache(). All cache-specific functions were instrumented with a BUG_ON() to make sure we never call them with cache disabled. Additionally the cache[] array was not initialized (remains NULL) so that we can later drop it if not needed. It's particularly huge and should be turned to dynamic with a pointer to a per-thread area where all the objects are located. This will solve the memory usage issue and will improve locality, or even help better deal with NUMA machines once each thread uses its own arena.	2022-02-23 17:11:33 +01:00
Willy Tarreau	dff3b0627d	MINOR: pools: make the global pools a runtime option. There were very few functions left that were specific to global pools, and even the checks they used to participate to are not directly on the most critical path so they can suffer an extra "if". What's done now is that pool_releasable() always returns 0 when global pools are disabled (like the one before) so that pool_evict_last_items() never tries to place evicted objects there. As such there will never be any object in the free list. However pool_refill_local_from_shared() is bypassed when global pools are disabled so that we even avoid the atomic loads from this function. The default global setting is still adjusted based on the original CONFIG_NO_GLOBAL_POOLS that is set depending on threads and the allocator. The global executable only grew by 1.1kB by keeping this code enabled, and the code is simplified and will later support runtime options.	2022-02-23 17:11:33 +01:00
Willy Tarreau	6f3c7f6e6a	MINOR: pools: add a new debugging flag POOL_DBG_INTEGRITY The test to decide whether or not to enforce integrity checks on cached objects is now enabled at runtime and conditionned by this new debugging flag. While previously it was not a concern to inflate the code size by keeping the two functions static, they were moved to pool.c to limit the impact. In pool_get_from_cache(), the fast code path remains fast by having both flags tested at once to open a slower branch when either POOL_DBG_COLD_FIRST or POOL_DBG_INTEGRITY are set.	2022-02-23 17:11:33 +01:00
Willy Tarreau	d3470e1ce8	MINOR: pools: add a new debugging flag POOL_DBG_COLD_FIRST When enabling pools integrity checks, we usually prefer to allocate cold objects first in order to maximize the time the objects spend in the cache. In order to make this configurable at runtime, let's introduce a new debugging flag to control this allocation order. It is currently preset by the DEBUG_POOL_INTEGRITY build-time setting.	2022-02-23 17:11:33 +01:00
Willy Tarreau	fd8b737e2c	MINOR: pools: switch DEBUG_DONT_SHARE_POOLS to runtime This test used to appear at a single location in create_pool() to enable a check on the pool name or unconditionally merge similarly sized pools. This patch introduces POOL_DBG_DONT_MERGE and conditions the test on this new runtime flag, that is preset according to the aforementioned debugging option.	2022-02-23 17:11:33 +01:00
Willy Tarreau	8d0273ed88	MINOR: pools: switch the fail-alloc test to runtime only The fail-alloc test used to be enabled/disabled at build time using the DEBUG_FAIL_ALLOC macro, but it happens that the cost of the test is quite cheap and that it can be enabled as one of the pool_debugging options. This patch thus introduces the first POOL_DBG_FAIL_ALLOC option, whose default value depends on DEBUG_FAIL_ALLOC. The mem_should_fail() function is now always built, but it was made static since it's never used outside.	2022-02-23 17:11:33 +01:00
Willy Tarreau	605629b008	MINOR: pools: introduce a new pool_debugging global variable This read-mostly variable will be used at runtime to enable/disable certain pool-debugging features and will be set by the command-line parser. A future option -dP will take a number of debugging features as arguments to configure this variable's contents.	2022-02-23 17:11:33 +01:00
Willy Tarreau	b61fccdc3f	CLEANUP: init: remove the ifdef on HAPROXY_MEMMAX It's ugly, let's move it to defaults.h with all other ones and preset it to zero if not defined.	2022-02-23 17:11:33 +01:00
William Lallemand	b4a4ef6a29	MINOR: httpclient/lua: ability to set a server timeout Add the ability to set a "server timeout" on the httpclient with either the httpclient_set_timeout() API or the timeout argument in a request. Issue #1470.	2022-02-23 15:11:11 +01:00
Willy Tarreau	9de8a2b854	CLEANUP: pools: remove the now unused pool_is_crowded() This function was renderred obsolete by commit `a0b5831ee` ("MEDIUM: pools: centralize cache eviction in a common function") which replaced its last call inside the loop with a single call out of the loop to pool_releasable() as introduced by commit `91a8e28f9` ("MINOR: pool: add a function to estimate how many may be released at once"). Let's remove it before it becomes wrong and used again.	2022-02-21 20:44:26 +01:00
Christopher Faulet	dc523e3b89	BUG/MEDIUM: htx: Be sure to have a buffer to perform a raw copy of a message In htx_copy_msg(), if the destination buffer is empty, we perform a raw copy of the message instead of a copy block per block. But we must be sure the destianation buffer was really allocated. In other word, to perform a raw copy, the HTX message must be empty _AND_ it must have some free space available. This function is only used to copy an HTTP reply (for instance, an error or a redirect) in the buffer of the response channel. For now, we are sure the buffer was allocated because it is a pre-requisite to call stream analyzers. However, it may be a source of bug in future. This patch may be backported as far as 2.3.	2022-02-21 16:05:47 +01:00
Willy Tarreau	d439a49655	DEBUG: buffer: check in __b_put_blk() whether the buffer room is respected This adds a BUG_ON() to make sure we don't face other situations like the one fixed by previous commit.	2022-02-18 17:33:27 +01:00
William Lallemand	7b2e0ee1c1	MINOR: httpclient: sets an alternative destination httpclient_set_dst() allows to set an alternative destination address using HAProxy addres format. This will ignore the address within the URL.	2022-02-17 20:07:00 +01:00
Frédéric Lécaille	71f3abbb52	MINOR: quic: Move quic_rxbuf_pool pool out of xprt part This pool could be confuse with that of the RX buffer pool for the connection (quic_conn_rxbuf).	2022-02-15 17:33:21 +01:00
Frédéric Lécaille	eca47d9a8a	MINOR: quic: Wrong smoothed rtt initialization In ->srtt we store 8srtt to ease the srtt computations with this formula: srtt = 7/8 srtt + 1/8 * adjusted_rtt But its initialization was wrong.	2022-02-15 17:23:44 +01:00
Amaury Denoyelle	9a327a7c3f	MINOR: mux-quic: implement rcv_buf Implement the stream rcv_buf operation on QUIC mux. A new buffer is stored in qcs structure named app_buf. This new buffer will contains HTX and will be filled for example on H3 DATA frame parsing. The rcv_buf operation transfer as much as possible data from the HTX from app_buf to the conn-stream buffer. This is mainly identical to mux-h2. This is required to support HTTP POST data.	2022-02-15 17:10:51 +01:00
Amaury Denoyelle	8524f0f779	MINOR: quic: use a global dghlrs for each thread Move the QUIC datagram handlers oustide of the receivers. Use a global handler per-thread which is allocated on post-config. Implement a free function on process deinit to avoid a memory leak.	2022-02-15 10:13:20 +01:00
Willy Tarreau	6c8babf6c4	BUG/MAJOR: sched: prevent rare concurrent wakeup of multi-threaded tasks Since the relaxation of the run-queue locks in 2.0 there has been a very small but existing race between expired tasks and running tasks: a task might be expiring and being woken up at the same time, on different threads. This is protected against via the TASK_QUEUED and TASK_RUNNING flags, but just after the task finishes executing, it releases it TASK_RUNNING bit an only then it may go to task_queue(). This one will do nothing if the task's ->expire field is zero, but if the field turns to zero between this test and the call to __task_queue() then three things may happen: - the task may remain in the WQ until the 24 next days if it's in the future; - the task may prevent any other task after it from expiring during the 24 next days once it's queued - if DEBUG_STRICT is set on 2.4 and above, an abort may happen - since 2.2, if the task got killed in between, then we may even requeue a freed task, causing random behaviour next time it's found there, or possibly corrupting the tree if it gets reinserted later. The peers code is one call path that easily reproduces the case with the ->expire field being reset, because it starts by setting it to TICK_ETERNITY as the first thing when entering the task handler. But other code parts also use multi-threaded tasks and rightfully expect to be able to touch their expire field without causing trouble. No trivial code path was found that would destroy such a shared task at runtime, which already limits the risks. This must be backported to 2.0.	2022-02-14 20:10:43 +01:00
Willy Tarreau	27c8da1fd5	DEBUG: pools: replace the link pointer with the caller's address on pool_free() Along recent evolutions of the pools, we've lost the ability to reliably detect double-frees because while in the past the same pointer was being used to chain the objects in the cache and to store the pool's address, since 2.0 they're different so the pool's address is never overwritten on free() and a double-free will rarely be detected. This patch sets the caller's return address there. It can never be equal to a pool's address and will help guess what was the previous call path. It will not work on exotic architectures nor with very old compilers but these are not the environments where we're trying to get detailed bug reports, and this is not done by default anyway so we don't care about this limitation. Note that depending on the inlining status of the function, the result may differ but that's no big deal either. A test by placing a double free of an appctx inside the release handler itself successfully reported the trouble during appctx_free() and showed that the return address was in stream_int_shutw_applet() (this one calls the release handler).	2022-02-14 20:10:43 +01:00
Willy Tarreau	49bb5d4268	DEBUG: pools: let's add reverse mapping from cache heads to thread and pool During global eviction we're visiting nodes from the LRU tail and we determine their pool cache head and their pool. In order to make sure we never mess up, let's add some backwards pointer to the thread number and pool from the pool_cache_head. It's 64-byte aligned anyway so we're not wasting space and it helps for debugging and will prevent memory corruption the earliest possible.	2022-02-14 20:10:43 +01:00
Willy Tarreau	e2830addda	DEBUG: pools: add extra sanity checks when picking objects from a local cache These few checks are added to make sure we never try to pick an object from an empty list, which would have a devastating effect.	2022-02-14 20:10:43 +01:00
Willy Tarreau	c895c441c7	BUG/MINOR: pools: always flush pools about to be destroyed When destroying a pool (e.g. at exit or when resizing buffers), it's important to try to free all their local objects otherwise we can leave some in the cache. This is particularly visible when changing "bufsize", because "show pools" will then show two "trash" pools, one of which contains a single object in cache (which is fortunately not reachable). In all cases this happens while single-threaded so that's easy to do, we just have to do it on the current thread. The easiest way to do this is to pass an extra argument to function pool_evict_from_local_cache() to force a full flush instead of a partial one. This can probably be backported to about all branches where this applies, but at least 2.4 needs it.	2022-02-14 20:10:43 +01:00
Frédéric Lécaille	83cd51e87a	MINOR: quic: Remove an RX buffer useless lock This lock is no more useful: the RX buffer for a connection is always handled by the same thread.	2022-02-14 15:20:54 +01:00
Remi Tricot-Le Breton	c76c3c4e59	MEDIUM: ssl: Replace all DH objects by EVP_PKEY on OpenSSLv3 (via HASSL_DH type) DH structure is a low-level one that should not be used anymore with OpenSSLv3. All functions working on DH were marked as deprecated and this patch replaces the ones we used with new APIs recommended in OpenSSLv3, be it in the migration guide or the multiple new manpages they created. This patch replaces all mentions of the DH type by the HASSL_DH one, which will be replaced by EVP_PKEY with OpenSSLv3 and will remain DH on older versions. It also uses all the newly created helper functions that enable for instance to load DH parameters from a file into an EVP_PKEY, or to set DH parameters into an SSL_CTX for use in a DHE negotiation. The following deprecated functions will effectively disappear when building with OpenSSLv3 : DH_set0_pqg, PEM_read_bio_DHparams, DH_new, DH_free, DH_up_ref, SSL_CTX_set_tmp_dh.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	55d7e782ee	MINOR: ssl: Set default dh size to 2048 Starting from OpenSSLv3, we won't rely on the SSL_CTX_set_tmp_dh_callback mechanism so we will need to know the DH size we want to use during init. In order for the default DH param size to be used when no RSA or DSA private key can be found for a given bind line, we will need to know the default size we want to use (which was not possible the way the code was built, since the global default dh size was set too late.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	09ebb3359a	MINOR: ssl: Add ssl_sock_get_dh_from_bio helper function This new function makes use of the new OpenSSLv3 APIs that should be used to load DH parameters from a file (or a BIO in this case) and that should replace the deprecated PEM_read_bio_DHparams function. Note that this function returns an EVP_PKEY when using OpenSSLv3 since they now advise against using low level structures such as DH ones. This helper function is not used yet so this commit should be stricly iso-functional, regardless of the OpenSSL version.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	956f3aea03	MINOR: ssl: Create HASSL_DH wrapper structure The DH mechanism relies on DH objects that are low-level structures that should not be used anymore starting from OpenSSLv3. With the newer OpenSSL version, we should only use higher level EVP_PKEY objects. Since enforcing this new logic to older versions of OpenSSL could be dangerous (or plain impossible), we will keeptwo versions of the code when required. The HASSL_DH define will allow to unify some of the functions that were created for DH use without having to add too many duplicated blocks of code depending on the OpenSSL version.	2022-02-14 10:07:14 +01:00
Remi Tricot-Le Breton	1effd9aa09	MINOR: ssl: Remove call to ERR_func_error_string with OpenSSLv3 ERR_func_error_string does not return anything anymore with OpenSSLv3, it can be replaced by ERR_peek_error_func which did not exist on previous versions.	2022-02-14 10:07:14 +01:00
Amaury Denoyelle	58a7704d54	MINOR: quic: take out xprt snd_buf operation Rename quic_conn_to_buf to qc_snd_buf and remove it from xprt ops. This is done to reflect the true usage of this function which is only a wrapper around sendto but cannot be called by the upper layer. qc_snd_buf is moved in quic-sock because to mark its link with quic_sock_fd_iocb which is the recvfrom counterpart.	2022-02-09 15:57:46 +01:00
Amaury Denoyelle	59e0b1f44c	MINOR: mux-quic: remove quic_transport_params_update This function is unused.	2022-02-09 15:05:23 +01:00
Amaury Denoyelle	b78805488b	MINOR: h3: hardcode the stream id of control stream Use the value of 0x3 for the stream-id of H3 control-stream. As a consequence, qcs_get_next_id is now unused and is thus removed.	2022-02-09 15:05:23 +01:00
Remi Tricot-Le Breton	8ea1f5f6cd	MINOR: ssl: Remove call to SSL_CTX_set_tlsext_ticket_key_cb with OpenSSLv3 SSL_CTX_set_tlsext_ticket_key_cb was deprecated on OpenSSLv3 because it uses an HMAC_pointer which is deprecated as well. According to the v3's manpage it should be replaced by SSL_CTX_set_tlsext_ticket_key_evp_cb which uses a EVP_MAC_CTX pointer. This new callback was introduced in OpenSSLv3 so we need to keep the two calls in the source base and to split the usage depending on the OpenSSL version.	2022-02-09 12:11:31 +01:00
Remi Tricot-Le Breton	36f80f6e0b	CLEANUP: ssl: Remove unused ssl_sock_create_cert function This function is not used anymore, it can be removed.	2022-02-09 11:15:44 +01:00
Ilya Shipitsin	e5ea76a3df	BUILD: ssl: adjust guard for X509_get_X509_PUBKEY(x) BoringSSL defines that function since https://boringssl.googlesource.com/boringssl/+/33f8d33af0dcb083610e978baad5a8b6e1cfee82	2022-02-02 17:47:57 +01:00
William Lallemand	2a17191e91	MINOR: mworker/cli: mcli-debug-mode enables every command "mcli-debug-mode on" enables every command that were meant for a worker, on the CLI of the master. Which mean you can issue, "show fd", show stat" in order to debug the MASTER proxy. You can also combine it with "expert-mode on" or "experimental-mode on" to access to more commands.	2022-02-02 15:51:24 +01:00
William Lallemand	7267f78ebe	MINOR: mworker/cli: set expert/experimental mode from the CLI Allow to set the master CLI in expert or experimental mode. No command within the master are unlocked yet, but it gives the ability to send expert or experimental commands to the workers. echo "@1; experimental-mode on; del server be1/s2" \| socat /var/run/haproxy.master - echo "experimental-mode on; @1 del server be1/s2" \| socat /var/run/haproxy.master -	2022-02-01 17:33:06 +01:00
Willy Tarreau	08b6f96452	MINOR: listener: replace the listener's spinlock with an rwlock We'll need to lock the listener a little bit more during accept() and tests show that a spinlock is a massive performance killer, so let's first switch to an rwlock for this lock. This patch might have to be backported for the next patch to work, and if so, the change is almost mechanical (look for LISTENER_LOCK), but do not forget about the few HA_SPIN_INIT() in the file. There's no reference to this lock outside of listener.c nor listener-t.h.	2022-02-01 16:51:55 +01:00
Amaury Denoyelle	aebe26f8ba	MINOR: mux-quic: create a timeout task This task will be used to schedule a timer when there is no activity on the mux. The timeout is set via the "timeout client" from the configuration file. The timeout task process schedule the timeout only on specific conditions. Currently, it's done if there is no opened bidirectional stream. For now this task is not used. This will be implemented in the following commit.	2022-02-01 15:19:35 +01:00
Willy Tarreau	9aa324de2d	DEBUG: fd: make sure we never try to insert/delete an impossible FD number It's among the cases that would provoke memory corruption, let's add some tests against negative FDs and those larger than the table. This must never ever happen and would currently result in silent corruption or a crash. Better have a noticeable one exhibiting the call chain if that were to happen.	2022-01-31 21:00:35 +01:00
Frédéric Lécaille	91f083a365	MINOR: quic: Do not try to accept a connection more than one time We add a new flag to mark a connection as already enqueued for acception. This is useful for 0-RTT session where a connection is first enqueued for acception as soon as 0-RTT RX secrets could be derived. Then as for any other connection, we could accept one more time this connection after handshake completion which lead to very bad side effects. Thank you to Amaury for this nice patch.	2022-01-31 16:40:23 +01:00
William Lallemand	56be0e0146	MINOR: mworker: allocate and initialize a mworker_proc mworker_proc_new() allocates and initializes correctly a mworker_proc structure.	2022-01-28 23:52:36 +01:00
William Lallemand	55a921c914	BUG/MINOR: mworker: fix a FD leak of a sockpair upon a failed reload When starting HAProxy in master-worker, the master pre-allocate a struct mworker_proc and do a socketpair() before the configuration parsing. If the configuration loading failed, the FD are never closed because they aren't part of listener, they are not even in the fdtab. This patch fixes the issue by cleaning the mworker_proc structure that were not asssigned a process, and closing its FDs. Must be backported as far as 2.0, the srv_drop() only frees the memory and could be dropped since it's done before an exec().	2022-01-28 23:47:43 +01:00
Willy Tarreau	cc5cd5b8d8	BUILD: task: use list_to_mt_list() instead of casting list to mt_list There were a few casts of list* to mt_list* that were upsetting some old compilers (not sure about the effect on others). We had created list_to_mt_list() purposely for this, let's use it instead of applying this cast.	2022-01-28 19:04:02 +01:00
Willy Tarreau	8f0b4e97e7	BUILD: tree-wide: mark a few numeric constants as explicitly long long At a few places in the code the switch/case ond flags are tested against 64-bit constants without explicitly being marked as long long. Some 32-bit compilers complain that the constant is too large for a long, and other likely always use long long there. Better fix that as it's uncertain what others which do not complain do. It may be backported to avoid doubts on uncommon platforms if needed, as it touches very few areas.	2022-01-28 19:04:02 +01:00
Willy Tarreau	95d3eaff36	BUILD: checks: fix inlining issue on set_srv_agent_[addr,port} These functions are declared as external functions in check.h and as inline functions in check.c. Let's move them as static inline in check.h. This appeared in 2.4 with the following commits: `4858fb2e1` ("MEDIUM: check: align agentaddr and agentport behaviour") `1c921cd74` ("BUG/MINOR: check: consitent way to set agentaddr") While harmless (it only triggers build warnings with some gcc 4.x), it should probably be backported where the paches above are present to keep the code consistent.	2022-01-28 19:04:02 +01:00
Willy Tarreau	a65b4933ba	BUILD: cpuset: do not use const on the source of CPU_AND/CPU_ASSIGN The man page indicates that CPU_AND() and CPU_ASSIGN() take a variable, not a const on the source, even though it doesn't make much sense. But with older libcs, this triggers a build warning: src/cpuset.c: In function 'ha_cpuset_and': src/cpuset.c:53: warning: initialization discards qualifiers from pointer target type src/cpuset.c: In function 'ha_cpuset_assign': src/cpuset.c:101: warning: initialization discards qualifiers from pointer target type Better stick stricter to the documented API as this is really harmless here. There's no need to backport it (unless build issues are reported, which is quite unlikely).	2022-01-28 19:04:02 +01:00
Willy Tarreau	8da23393a1	BUILD: atomic: make the old HA_ATOMIC_LOAD() support const pointers We have an implementation of atomic ops for older versions of gcc that do not provide the __builtin_* API (< 4.4). Recent changes to the pools broke that in pool_releasable() by having a load from a const pointer, which doesn't work there due to a temporary local variable that is declared then assigned. Let's make use of a compount statement to assign it a value when declaring it. There's no need to backport this.	2022-01-28 19:04:02 +01:00
Willy Tarreau	b510116fd2	MINOR: sock: move the unused socket cleaning code into its own function The startup code used to scan the list of unused sockets retrieved from an older process, and to close them one by one. This also required that the knowledge of the internal storage of these temporary sockets was known from outside sock.c and that the code was copy-pasted at every call place. This patch moves this into sock.c under the name sock_drop_unused_old_sockets(), and removes the xfer_sock_list definition from sock.h since the rest of the code doesn't need to know this. This cleanup is minimal and preliminary to a future fix that will need to be backported to all versions featuring FD transfers over the CLI.	2022-01-28 19:04:02 +01:00
Amaury Denoyelle	0442efd214	MINOR: quic: refactor quic CID association with threads Do not use an extra DCID parameter on new_quic_cid to be able to associated a new generated CID to a thread ID. Simply do the computation inside the function. The API is cleaner this way. This also has the effects to improve the apparent randomness of CIDs. With the previous version the first byte of all CIDs are identical for a connection which could lead to privacy issue. This version may not be totally perfect on this aspect but it improves the situation.	2022-01-28 16:29:27 +01:00
Frédéric Lécaille	dc36404c36	MINOR: quic: Drop Initial packets with wrong ODCID According to the RFC 9000, the client ODCID must have a minimal length of 8 bytes.	2022-01-28 16:08:07 +01:00
Frédéric Lécaille	74904a4792	MINOR: quic: Make usage of by datagram handler trees The CID trees are no more attached to the listener receiver but to the underlying datagram handlers (one by thread) which run always on the same thread. So, any operation on these trees do not require any locking.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	9ea9463d47	MINOR: quic: Attach all the CIDs to the same connection We copy the first octet of the original destination connection ID to any CID for the connection calling new_quic_cid(). So this patch modifies only this function to take a dcid as passed parameter.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	37ae505c21	MINOR: quic: Do not consume the RX buffer on QUIC sock i/o handler side Rename quic_lstnr_dgram_read() to quic_lstnr_dgram_dispatch() to reflect its new role. After calling this latter, the sock i/o handler must consume the buffer only if the datagram it received is detected as wrong by quic_lstnr_dgram_dispatch(). The datagram handler task mark the datagram as consumed atomically setting ->buf to NULL value. The sock i/o handler is responsible of flushing its RX buffer before using it. It also keeps a datagram among the consumed ones so that to pass it to quic_lstnr_dgram_dispatch() and prevent it from allocating a new one.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	794d068d8f	MINOR: proto_quic: Wrong allocations for TX rings and RX bufs As mentionned in the comment, the tx_qrings and rxbufs members of receiver struct must be pointers to pointers! Modify the functions responsible of their allocations consequently. Note that this code could work because sizeof rxbuf and sizeof tx_qrings are greater than the size of pointer!	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	d152309423	CLEANUP: quic: Remove useless definition The quic_dgram_ctx struct has been replaced by quic_dgram struct. There is no need to keek a typedef for a pointer to function since we converted the UDP datagram parser (quic_dgram_read()) into a task.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	25bc8875d7	MINOR: quic: Convert quic_dgram_read() into a task quic_dgram_read() parses all the QUIC packets from a UDP datagram. It is the best candidate to be converted into a task, because is processing data unit is the UDP datagram received by the QUIC sock i/o handler. If correct, this datagram is added to the context of a task, quic_lstnr_dghdlr(), a conversion of quic_dgram_read() into a task. This task pop a datagram from an mt_list and passes it among to the packet handler (quic_lstnr_pkt_rcv()). Modify the quic_dgram struct to play the role of the old quic_dgram_ctx struct when passed to quic_lstnr_pkt_rcv(). Modify the datagram handlers allocation to set their tasks to quic_lstnr_dghdlr().	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	220894a5d6	MINOR: quic: Pass CID as a buffer to quic_get_cid_tid() Very minor modification so that this function might be used for a context without CID (at datagram level).	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	69dd5e6a0b	MINOR: proto_quic: Allocate datagram handlers Add quic_dghdlr new struct do define datagram handler tasks, one by thread. Allocate them and attach them to the listener receiver part calling quic_alloc_dghdlrs_listener() newly implemented function.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	3d4bfe708a	MINOR: quic: Allocate QUIC datagrams from sock I/O handler Add quic_dgram new structure to store information about datagrams received by the sock I/O handler (quic_sock_fd_iocb) and its associated pool. Implement quic_get_dgram_dcid() to retrieve the datagram DCID which must be the same for all the packets in the datagram. Modify quic_lstnr_dgram_read() called by the sock I/O handler to allocate a quic_dgram each time a correct datagram is found and add it to the sock I/O handler rxbuf dgram list.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	53898bba81	MINOR: quic: Add a list to QUIC sock I/O handler RX buffer This list will be used to store datagrams in the rxbuf struct used by the quic_sock_fd_iocb() QUIC sock I/O handler with one rxbuf by thread.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	ce521e4f15	MINOR: quic: Add new defintion about DCIDs offsets Define the offsets of the DCIDs from the beginning of a QUIC packets. Note that they must always be present. As QUIC servers, QUIC haproxy listeners always use a CID, source CID on the haproxy side, which is a destination ID on the peer side.	2022-01-27 16:37:55 +01:00
Frédéric Lécaille	3d55462654	MINOR: quic: Get rid of a struct buffer in quic_lstnr_dgram_read() This is to be sure xprt functions do not manipulate the buffer struct passed as parameter to quic_lstnr_dgram_read() from low level datagram I/O callback in quic_sock.c (quic_sock_fd_iocb()).	2022-01-27 16:37:55 +01:00
Willy Tarreau	ecc473b529	BUG/MAJOR: compiler: relax alignment constraints on certain structures In github bug #1517, Mike Lothian reported instant crashes on startup on RHEL8 + gcc-11 that appeared with 2.4 when allocating a proxy. The analysis brought us down to the THREAD_ALIGN() entries that were placed inside the "server" struct to avoid false sharing of cache lines. It turns out that some modern gcc make use of aligned vector operations to manipulate some fields (e.g. memset() etc) and that these structures allocated using malloc() are not necessarily aligned, hence the crash. The compiler is allowed to do that because the structure claims to be aligned. The problem is in fact that the alignment propagates to other structures that embed it. While most of these structures are used as statically allocated variables, some are dynamic and cannot use that. A deeper analysis showed that struct server does this, propagates to struct proxy, which propagates to struct spoe_config, all of which are allocated using malloc/calloc. A better approach would consist in usins posix_memalign(), but this one is not available everywhere and will either need to be reimplemented less efficiently (by always wasting 64 bytes before the area), or a few functions will have to be specifically written to deal with the few structures that are dynamically allocated. But the deeper problem that remains is that it is difficult to track structure alignment, as there's no available warning to check this. For the long term we'll probably have to create a macro such as "struct_malloc()" etc which takes a type and enforces an alignment based on the one of this type. This also means propagating that to pools as well, and it's not a tiny task. For now, let's get rid of the forced alignment in struct server, and replace it with extra padding. By punching 63-byte holes, we can keep areas on separate cache lines. Doing so moderately increases the size of the "server" structure (~+6%) but that's the best short-term option and it's easily backportable. This will have to be backported as far as 2.4. Thanks to Mike for the detailed report.	2022-01-27 16:28:10 +01:00
Amaury Denoyelle	cfa2d5648f	MAJOR: quic: implement accept queue Do not proceed to direct accept when creating a new quic_conn. Wait for the QUIC handshake to succeeds to insert the quic_conn in the accept queue. A tasklet is then woken up to call listener_accept to accept the quic_conn. The most important effect is that the connection/mux layers are not instantiated at the same time as the quic_conn. This forces to delay some process to be sure that the mux is allocated : * initialization of mux transport parameters * installation of the app-ops Also, the mux instance is not checked now to wake up the quic_conn tasklet. This is safe because the xprt-quic code is now ready to handle the absence of the connection/mux layers. Note that this commit has a deep impact as it changes significantly the lower QUIC architecture. Most notably, it breaks the 0-RTT feature.	2022-01-26 16:13:54 +01:00
Amaury Denoyelle	f68b2cb816	MINOR: listener: define per-thr struct Create a new structure li_per_thread. This is uses as an array in the listener structure, with an entry allocated per thread. The new function li_init_per_thr is responsible of the allocation. For now, li_per_thread contains fields only useful for QUIC listeners. As such, it is only allocated for QUIC listeners.	2022-01-26 16:13:54 +01:00
Amaury Denoyelle	2ce99fe4bf	MINOR: quic: create accept queue for QUIC connections Create a new type quic_accept_queue to handle QUIC connections accept. A queue will be allocated for each thread. It contains a list of listeners which contains at least one quic_conn ready to be accepted and the tasklet to run listener_accept for these listeners.	2022-01-26 16:13:51 +01:00
Amaury Denoyelle	b59b88950a	MINOR: quic: define QUIC flag on listener Mark QUIC listeners with the flag LI_F_QUIC_LISTENER. It is set by the proto-quic layer on the add listener callback. This allows to override more clearly the accept callback on quic_session_accept.	2022-01-26 15:25:45 +01:00
Amaury Denoyelle	31ea9177ac	MINOR: listener: add flags field Define a new field in listener structure named flags. For the moment, no flag is defined. This will be notably useful to differentiate QUIC listeners with the implementation of a QUIC conn accept queue.	2022-01-26 15:25:45 +01:00
Amaury Denoyelle	9fa15e5413	MINOR: quic: do not manage connection in xprt snd_buf Remove usage of connection in quic_conn_from_buf. As connection and quic_conn are decorrelated, it is not logical to check connection flags when using sendto. This require to store the L4 peer address in quic_conn to be able to use sendto. This change is required to delay allocation of connection.	2022-01-26 15:25:38 +01:00
Amaury Denoyelle	7f7713d6ef	MINOR: receiver: define a flag for local accept This flag is named RX_F_LOCAL_ACCEPT. It will be activated for special receivers where connection balancing to threads is already handle outside of listener_accept, such as with QUIC listeners.	2022-01-26 11:22:20 +01:00
Amaury Denoyelle	4b40f19f92	MINOR: quic: refactor app-ops initialization Add a new function in mux-quic to install app-ops. For now this functions is called during the ALPN negotiation of the QUIC handshake. This change will be useful when the connection accept queue will be implemented. It will be thus required to delay the app-ops initialization because the mux won't be allocated anymore during the QUIC handshake.	2022-01-26 10:59:33 +01:00
Amaury Denoyelle	0b1f93127f	MINOR: quic: handle app data according to mux/connection layer status Define a new enum to represent the status of the mux/connection layer above a quic_conn. This is important to know if it's possible to handle application data, or if it should be buffered or dropped.	2022-01-26 10:57:17 +01:00
Willy Tarreau	add43fa43e	DEBUG: pools: add new build option DEBUG_POOL_TRACING This new option, when set, will cause the callers of pool_alloc() and pool_free() to be recorded into an extra area in the pool that is expected to be helpful for later inspection (e.g. in core dumps). For example it may help figure that an object was released to a pool with some sub-fields not yet released or that a use-after-free happened after releasing it, with an immediate indication about the exact line of code that released it (possibly an error path). This only works with the per-thread cache, and even objects refilled from the shared pool directly into the thread-local cache will have a NULL there. That's not an issue since these objects have not yet been freed. It's worth noting that pool_alloc_nocache() continues not to set any caller pointer (e.g. when the cache is empty) because that would require a possibly undesirable API change. The extra cost is minimal (one pointer per object) and this completes well with DEBUG_POOL_INTEGRITY.	2022-01-24 16:40:48 +01:00
Willy Tarreau	0e2a5b4b61	MINOR: pools: extend pool_cache API to pass a pointer to a caller This adds a caller to pool_put_to_cache() and pool_get_from_cache() which will optionally be used to pass a pointer to their callers. For now it's not used, only the API is extended to support this pointer.	2022-01-24 16:40:48 +01:00
Willy Tarreau	7fa092b727	MINOR: pools: prepare POOL_EXTRA to be split into multiple extra fields Here the idea is to calculate the POOL_EXTRA size that is appended at the end of a pool object based on the sum of enabled optional fields so that we can more easily compute offsets and sizes depending on build options. For this, POOL_EXTRA is replaced with POOL_EXTRA_MARK which itself is set either to sizeof(void*) or zero depending on whether we enable marking the origin pool or not upon allocation.	2022-01-24 16:40:48 +01:00
Willy Tarreau	d392973dcc	MINOR: pools: partially uninline pool_alloc() The pool_alloc() function was already a wrapper to __pool_alloc() which was also inlined but took a set of flags. This latter was uninlined and moved to pool.c, and pool_alloc()/pool_zalloc() turned to macros so that they can more easily evolve to support debugging options. The number of call places made this code grow over time and doing only this change saved ~1% of the whole executable's size.	2022-01-24 16:40:48 +01:00
Willy Tarreau	15c322c413	MINOR: pools: partially uninline pool_free() The pool_free() function has become a bit big over time due to the extra consistency checks. It used to remain inline only to deal cleanly with the NULL pointer free that's quite present on some structures (e.g. in stream_free()). Here we're splitting the function in two: - __pool_free() does the inner block without the pointer test and becomes a function ; - pool_free() is now a macro that only checks the pointer and calls __pool_free() if needed. The use of a macro versus an inline function is only motivated by an easier intrumentation of the code later. With this change, the code size reduces by ~1%, which means that at this point all pool_free() call places used to represent more than 1% of the total code size.	2022-01-24 16:40:48 +01:00
Amaury Denoyelle	9320dd5385	MEDIUM: quic/ssl: add new ex data for quic_conn Allow to register quic_conn as ex-data in SSL callbacks. A new index is used to identify it as ssl_qc_app_data_index. Replace connection by quic_conn as SSL ex-data when initializing the QUIC SSL session. When using SSL callbacks in QUIC context, the connection is now NULL. Used quic_conn instead to retrieve the required parameters. Also clean up The same changes are conducted inside the QUIC SSL methods of xprt-quic : connection instance usage is replaced by quic_conn.	2022-01-24 10:30:49 +01:00
Amaury Denoyelle	57af069571	MINOR: quic: set listener accept cb on parsing Define a special accept cb for QUIC listeners to quic_session_accept(). This operation is conducted during the proto.add callback when creating listeners. A special care is now taken care when setting the standard callback session_accept_fd() to not overwrite if already defined by the proto layer.	2022-01-24 10:30:49 +01:00
Willy Tarreau	0575d8fd76	DEBUG: pools: add new build option DEBUG_POOL_INTEGRITY When enabled, objects picked from the cache are checked for corruption by comparing their contents against a pattern that was placed when they were inserted into the cache. Objects are also allocated in the reverse order, from the oldest one to the most recent, so as to maximize the ability to detect such a corruption. The goal is to detect writes after free (or possibly hardware memory corruptions). Contrary to DEBUG_UAF this cannot detect reads after free, but may possibly detect later corruptions and will not consume extra memory. The CPU usage will increase a bit due to the cost of filling/checking the area and for the preference for cold cache instead of hot cache, though not as much as with DEBUG_UAF. This option is meant to be usable in production.	2022-01-21 19:07:48 +01:00
Willy Tarreau	6c539c4b8c	BUG/MINOR: stream: make the call_rate only count the no-progress calls We have an anti-looping protection in process_stream() that detects bugs that used to affect a few filters like compression in the past which sometimes forgot to handle a read0 or a particular error, leaving a thread looping at 100% CPU forever. When such a condition is detected, an alert it emitted and the process is killed so that it can be replaced by a sane one: [ALERT] (19061) : A bogus STREAM [0x274abe0] is spinning at 2057156 calls per second and refuses to die, aborting now! Please report this error to developers [strm=0x274abe0,3 src=unix fe=MASTER be=MASTER dst=<MCLI> txn=(nil),0 txn.req=-,0 txn.rsp=-,0 rqf=c02000 rqa=10000 rpf=88000021 rpa=8000000 sif=EST,40008 sib=DIS,84018 af=(nil),0 csf=0x274ab90,8600 ab=0x272fd40,1 csb=(nil),0 cof=0x25d5d80,1300:PASS(0x274aaf0)/RAW((nil))/unix_stream(9) cob=(nil),0:NONE((nil))/NONE((nil))/NONE(0) filters={}] call trace(11): \| 0x4dbaab [c7 04 25 01 00 00 00 00]: stream_dump_and_crash+0x17b/0x1b4 \| 0x4df31f [e9 bd c8 ff ff 49 83 7c]: process_stream+0x382f/0x53a3 (...) One problem with this detection is that it used to only count the call rate because we weren't sure how to make it more accurate, but the threshold was high enough to prevent accidental false positives. There is actually one case that manages to trigger it, which is when sending huge amounts of requests pipelined on the master CLI. Some short requests such as "show version" are sufficient to be handled extremely fast and to cause a wake up of an analyser to parse the next request, then an applet to handle it, back and forth. But this condition is not an error, since some data are being forwarded by the stream, and it's easy to detect it. This patch modifies the detection so that update_freq_ctr() only applies to calls made without CF_READ_PARTIAL nor CF_WRITE_PARTIAL set on any of the channels, which really indicates that nothing is happening at all. This is greatly sufficient and extremely effective, as the call above is still caught (shutr being ignored by an analyser) while a loop on the master CLI now has no effect. The "call_rate" field in the detailed "show sess" output will now be much lower, except for bogus streams, which may help spot them. This field is only there for developers anyway so it's pretty fine to slightly adjust its meaning. This patch could be backported to stable versions in case of reports of such an issue, but as that's unlikely, it's not really needed.	2022-01-20 18:56:57 +01:00
Frédéric Lécaille	82468ea98e	MINOR: quic: Remove the packet number space TX MT_LIST There is no need to use an MT_LIST to store frames to send from a packet number space. This is a reminiscence for multi-threading support for the TX part.	2022-01-20 16:43:06 +01:00
Willy Tarreau	c514365317	MINOR: channel: add new function co_getdelim() to support multiple delimiters For now we have co_getline() which reads a buffer and stops on LF, and co_getword() which reads a buffer and stops on one arbitrary delimiter. But sometimes we'd need to stop on a set of delimiters (CR and LF, etc). This patch adds a new function co_getdelim() which takes a set of delimiters as a string, and constructs a small map (32 bytes) that's looked up during parsing to stop after the first delimiter found within the set. It also supports an optional escape character that skips a delimiter (typically a backslash). For the rest it works exactly like the two other variants.	2022-01-19 19:16:47 +01:00
William Lallemand	bad9c8cac4	BUG/MINOR: httpclient: set default Accept and User-Agent headers Some servers require at least an Accept and a User-Agent header in the request. This patch sets some default value. Must be backported in 2.5.	2022-01-14 20:46:21 +01:00
Willy Tarreau	39fd546d4b	MINOR: pools: enable pools with DEBUG_FAIL_ALLOC as well During 2.4-dev, fault injection was enabled for cached pools with commit `207c09509` ("MINOR: pools: move the fault injector to __pool_alloc()"), except that the condition for CONFIG_HAP_POOLS still depended on DEBUG_FAIL_ALLOC not being set, which limits the usability to cases where the define is set by hand. Let's remove it from the equation as this is not a constraint anymore. While a bit old, there's no need to backport this as it's only used during development.	2022-01-12 17:31:01 +01:00
Amaury Denoyelle	b76ae69513	MEDIUM: quic: implement Retry emission Implement the emission of Retry packets. These packets are emitted in response to Initial from clients without token. The token from the Retry packet contains the ODCID from the Initial packet. By default, Retry packet emission is disabled and the handshake can continue without address validation. To enable Retry, a new bind option has been defined named "quic-force-retry". If set, the handshake must be conducted only after receiving a token in the Initial packet.	2022-01-12 11:08:48 +01:00

... 3 4 5 6 7 ...

6101 Commits