numam-dpdk

Author	SHA1	Message	Date
Bing Zhao	9847fd125d	ethdev: introduce conntrack flow action and item This commit introduces the conntrack action and item. Usually the HW offloading is stateless. For some stateful offloading like a TCP connection, HW module will help provide the ability of a full offloading w/o SW participation after the connection was established. The basic usage is that in the first flow rule the application should add the conntrack action and jump to the next flow table. In the following flow rule(s) of the next table, the application should use the conntrack item to match on the result. A TCP connection has two directions traffic. To set a conntrack action context correctly, the information of packets from both directions are required. The conntrack action should be created on one ethdev port and supply the peer ethdev port as a parameter to the action. After context created, it could only be used between these two ethdev ports (dual-port mode) or a single port. The application should modify the action via the API "rte_action_handle_update" only when before using it to create a flow rule with conntrack for the opposite direction. This will help the driver to recognize the direction of the flow to be created, especially in the single-port mode, in which case the traffic from both directions will go through the same ethdev port if the application works as an "forwarding engine" but not an end point. There is no need to call the update interface if the subsequent flow rules have nothing to be changed. Query will be supported via "rte_action_handle_query" interface, about the current packets information and connection status. The fields query capabilities depends on the HW. For the packets received during the conntrack setup, it is suggested to re-inject the packets in order to make sure the conntrack module works correctly without missing any packet. Only the valid packets should pass the conntrack, packets with invalid TCP information, like out of window, or with invalid header, like malformed, should not pass. Naming and definition: https://elixir.bootlin.com/linux/latest/source/include/uapi/linux/ netfilter/nf_conntrack_tcp.h https://elixir.bootlin.com/linux/latest/source/net/netfilter/ nf_conntrack_proto_tcp.c Other reference: https://www.usenix.org/legacy/events/sec01/invitedtalks/rooij.pdf Signed-off-by: Bing Zhao <bingz@nvidia.com> Acked-by: Ori Kam <orika@nvidia.com> Acked-by: Thomas Monjalon <thomas@monjalon.net>	2021-04-20 01:24:57 +02:00
Robin Zhang	e9c5672ac1	net/iavf: deprecate i40evf PMD The i40evf PMD will be deprecated, iavf will be the only VF driver for Intel 700 serial (i40e) NIC family. To reach this, there will be 2 steps: Step 1: iavf will be the default VF driver, while i40evf still can be selected by devarg: "driver=i40evf". This is covered by this patch, which include: 1) add all 700 serial NIC VF device ID into iavf PMD 2) skip probe if devargs contain "driver=i40evf" in iavf 3) continue probe if devargs contain "driver=i40evf" in i40evf Step 2: i40evf and related devarg are removed, this will happen at DPDK 21.11 Between step 1 and step 2, no new feature will be added into i40evf except bug fix. Signed-off-by: Robin Zhang <robinx.zhang@intel.com> Acked-by: Qi Zhang <qi.z.zhang@intel.com> Acked-by: Ferruh Yigit <ferruh.yigit@intel.com> Acked-by: Beilei Xing <beilei.xing@intel.com>	2021-04-19 10:36:17 +02:00
Ori Kam	b10a421a1f	ethdev: add packet integrity check flow rules Currently, DPDK application can offload the checksum check, and report it in the mbuf. However, as more and more applications are offloading some or all logic and action to the HW, there is a need to check the packet integrity so the right decision can be taken. The application logic can be positive meaning if the packet is valid jump / do actions, or negative if packet is not valid jump to SW / do actions (like drop) and add default flow (match all in low priority) that will direct the miss packet to the miss path. Since currently rte_flow works in positive way the assumption is that the positive way will be the common way in this case also. When thinking what is the best API to implement such feature, we need to consider the following (in no specific order): 1. API breakage. 2. Simplicity. 3. Performance. 4. HW capabilities. 5. rte_flow limitation. 6. Flexibility. First option: Add integrity flags to each of the items. For example add checksum_ok to IPv4 item. Pros: 1. No new rte_flow item. 2. Simple in the way that on each item the app can see what checks are available. Cons: 1. API breakage. 2. Increase number of flows, since app can't add global rule and must have dedicated flow for each of the flow combinations, for example matching on ICMP traffic or UDP/TCP traffic with IPv4 / IPv6 will result in 5 flows. Second option: dedicated item Pros: 1. No API breakage, and there will be no for some time due to having extra space. (by using bits) 2. Just one flow to support the ICMP or UDP/TCP traffic with IPv4 / IPv6. 3. Simplicity application can just look at one place to see all possible checks. 4. Allow future support for more tests. Cons: 1. New item, that holds number of fields from different items. For starter the following bits are suggested: 1. packet_ok - means that all HW checks depending on packet layer have passed. This may mean that in some HW such flow should be split to number of flows or fail. 2. l2_ok - all check for layer 2 have passed. 3. l3_ok - all check for layer 3 have passed. If packet doesn't have L3 layer this check should fail. 4. l4_ok - all check for layer 4 have passed. If packet doesn't have L4 layer this check should fail. 5. l2_crc_ok - the layer 2 CRC is O.K. 6. ipv4_csum_ok - IPv4 checksum is O.K. It is possible that the IPv4 checksum will be O.K. but the l3_ok will be 0. It is not possible that checksum will be 0 and the l3_ok will be 1. 7. l4_csum_ok - layer 4 checksum is O.K. 8. l3_len_OK - check that the reported layer 3 length is smaller than the frame length. Example of usage: 1. Check packets from all possible layers for integrity. flow create integrity spec packet_ok = 1 mask packet_ok = 1 ..... 2. Check only packet with layer 4 (UDP / TCP) flow create integrity spec l3_ok = 1, l4_ok = 1 mask l3_ok = 1 l4_ok = 1 Signed-off-by: Ori Kam <orika@nvidia.com> Acked-by: Ferruh Yigit <ferruh.yigit@intel.com> Acked-by: Ajit Khaparde <ajit.khaparde@broadcom.com> Acked-by: Thomas Monjalon <thomas@monjalon.net>	2021-04-19 19:05:17 +02:00
Bing Zhao	4b61b8774b	ethdev: introduce indirect flow action Right now, rte_flow_shared_action_* APIs are used for some shared actions, like RSS, count. The shared action should be created before using it inside a flow. These shared actions sometimes are not really shared but just some indirect actions decoupled from a flow. The new functions rte_flow_action_handle_* are added to replace the current shared functions rte_flow_shared_action_. There are two types of flow actions: 1. the direct (normal) actions that could be created and stored within a flow rule. Such action is tied to its flow rule and cannot be reused. 2. the indirect action, in the past, named shared_action. It is created from a direct actioni, like count or rss, and then used in the flow rules with an object handle. The PMD will take care of the retrieve from indirect action to the direct action when it is referenced. The indirect action is accessed (update / query) w/o any flow rule, just via the action object handle. For example, when querying or resetting a counter, it could be done out of any flow using this counter, but only the handle of the counter action object is required. The indirect action object could be shared by different flows or used by a single flow, depending on the direct action type and the real-life requirements. The handle of an indirect action object is opaque and defined in each driver and possibly different per direct action type. The old name "shared" is improper in a sense and should be replaced. Since the APIs are changed from "rte_flow_shared_action" to the new "rte_flow_action_handle", the testpmd application code and command line interfaces also need to be updated to do the adaption. The testpmd application user guide is also updated. All the "shared action" related parts are replaced with "indirect action" to have a correct explanation. The parameter of "update" interface is also changed. A general pointer will replace the rte_flow_action struct pointer due to the facts: 1. Some action may not support fields updating. In the example of a counter, the only "update" supported should be the reset. So passing a rte_flow_action struct pointer is meaningless and there is even no such corresponding action struct. What's more, if more than one operations should be supported, for some other action, such pointer parameter may not meet the need. 2. Some action may need conditional or partial update, the current parameter will not provide the ability to indicate which part(s) to update. For different types of indirect action objects, the pointer could either be the same of rte_flow_action struct - in order not to break the current driver implementation, or some wrapper structures with bits as masks to indicate which part to be updated, depending on real needs of the corresponding direct action. For different direct actions, the structures of indirect action objects updating will be different. All the underlayer PMD callbacks will be moved to these new APIs. The RTE_FLOW_ACTION_TYPE_SHARED is kept for now in order not to break the ABI. All the implementations are changed by using RTE_FLOW_ACTION_TYPE_INDIRECT. Since the APIs are changed from "rte_flow_shared_action" to the new "rte_flow_action_handle" and the "update" interface's 3rd input parameter is changed to generic pointer, the mlx5 PMD that uses these APIs needs to do the adaption to the new APIs as well. Signed-off-by: Bing Zhao <bingz@nvidia.com> Acked-by: Andrey Vesnovaty <andreyv@nvidia.com> Acked-by: Ori Kam <orika@nvidia.com> Acked-by: Ajit Khaparde <ajit.khaparde@broadcom.com> Acked-by: Thomas Monjalon <thomas@monjalon.net>	2021-04-19 18:25:42 +02:00
Lijun Ou	9ad9ff476c	ethdev: add queue state in queried queue information Currently, upper-layer application could get queue state only through pointers such as dev->data->tx_queue_state[queue_id], this is not the recommended way to access it. So this patch add get queue state when call rte_eth_rx_queue_info_get and rte_eth_tx_queue_info_get API. Note: After add queue_state field, the 'struct rte_eth_rxq_info' size remains 128B, and the 'struct rte_eth_txq_info' size remains 64B, so it could be ABI compatible. Signed-off-by: Chengwen Feng <fengchengwen@huawei.com> Signed-off-by: Lijun Ou <oulijun@huawei.com> Acked-by: Konstantin Ananyev <konstantin.ananyev@intel.com> Acked-by: Thomas Monjalon <thomas@monjalon.net> Reviewed-by: Ferruh Yigit <ferruh.yigit@intel.com>	2021-04-19 18:25:35 +02:00
Bruce Richardson	99a2dd955f	lib: remove librte_ prefix from directory names There is no reason for the DPDK libraries to all have 'librte_' prefix on the directory names. This prefix makes the directory names longer and also makes it awkward to add features referring to individual libraries in the build - should the lib names be specified with or without the prefix. Therefore, we can just remove the library prefix and use the library's unique name as the directory name, i.e. 'eal' rather than 'librte_eal' Signed-off-by: Bruce Richardson <bruce.richardson@intel.com>	2021-04-21 14:04:09 +02:00
Vladimir Medvedkin	28ebff11c2	hash: add predictable RSS This patch adds predictable RSS API. It is based on the idea of searching partial Toeplitz hash collisions. Signed-off-by: Vladimir Medvedkin <vladimir.medvedkin@intel.com> Acked-by: Konstantin Ananyev <konstantin.ananyev@intel.com> Acked-by: Yipeng Wang <yipeng1.wang@intel.com> Acked-by: John McNamara <john.mcnamara@intel.com>	2021-04-20 23:13:23 +02:00
Conor Walsh	6a094e3285	examples/l3fwd: implement FIB lookup method This patch implements the Forwarding Information Base (FIB) library in l3fwd using the function calls and infrastructure introduced in the previous patch. Signed-off-by: Conor Walsh <conor.walsh@intel.com> Acked-by: Konstantin Ananyev <konstantin.ananyev@intel.com> Acked-by: Vladimir Medvedkin <vladimir.medvedkin@intel.com>	2021-04-20 20:18:29 +02:00
Akhil Goyal	f96a8ebb27	eventdev: introduce crypto adapter enqueue API In case an event from a previous stage is required to be forwarded to a crypto adapter and PMD supports internal event port in crypto adapter, exposed via capability RTE_EVENT_CRYPTO_ADAPTER_CAP_INTERNAL_PORT_OP_FWD, we do not have a way to check in the API rte_event_enqueue_burst(), whether it is for crypto adapter or for eth tx adapter. Hence we need a new API similar to rte_event_eth_tx_adapter_enqueue(), which can send to a crypto adapter. Note that RTE_EVENT_TYPE_* cannot be used to make that decision, as it is meant for event source and not event destination. And event port designated for crypto adapter is designed to be used for OP_NEW mode. Hence, in order to support an event PMD which has an internal event port in crypto adapter (RTE_EVENT_CRYPTO_ADAPTER_OP_FORWARD mode), exposed via capability RTE_EVENT_CRYPTO_ADAPTER_CAP_INTERNAL_PORT_OP_FWD, application should use rte_event_crypto_adapter_enqueue() API to enqueue events. When internal port is not available(RTE_EVENT_CRYPTO_ADAPTER_OP_NEW mode), application can use API rte_event_enqueue_burst() as it was doing earlier, i.e. retrieve event port used by crypto adapter and bind its event queues to that port and enqueue events using the API rte_event_enqueue_burst(). Signed-off-by: Akhil Goyal <gakhil@marvell.com> Acked-by: Abhinandan Gujjar <abhinandan.gujjar@intel.com> Acked-by: Jerin Jacob <jerinj@marvell.com>	2021-04-17 18:49:52 +02:00
Yuying Zhang	2321e34c23	net/ice: support flow priority for DCF switch filter Support rte flow priority attribute for DCF switch filter. When a packet is matched by two rules, the behavior of it is not defined. This patch supports flow priority to create different recipes for this situation. Only priority 0 and 1 are supported and higher value denotes higher priority. for example: 1. flow create 0 priority 0 ingress pattern eth / vlan tci is 2 / vlan tci is 2 / end actions vf id 2 / end 2. flow create 0 priority 1 ingress pattern eth / vlan / vlan / ipv4 dst is 192.168.0.1 / end actions vf id 1 / end These two rules can be created at the same time in DCF switch filter and priority of rule 2 is higher. Packet hits rule 2 when two conditions of rules are satisfied. Signed-off-by: Yuying Zhang <yuying.zhang@intel.com> Acked-by: Qi Zhang <qi.z.zhang@intel.com>	2021-04-16 12:22:00 +02:00
Yuying Zhang	a65126d1ad	net/ice: support GTPU TEID pattern for switch filter Enable GTPU pattern for CVL switch filter. Support teid and qfi field of GTPU pattern. Patterns without inner l3/l4 field support outer dst/src ip. Patterns with inner l3/l4 field only support inner dst/src ip and inner dst/src port. +----------------------------------+------------------------------------+ \| Pattern \| Input Set \| +----------------------------------+------------------------------------+ \| pattern_eth_ipv4_gtpu \| teid, dst/src ip \| \| pattern_eth_ipv6_gtpu \| teid, dst/src ip \| \| pattern_eth_ipv4_gtpu_ipv4 \| teid, dst/src ip \| \| pattern_eth_ipv4_gtpu_ipv4_tcp \| teid, dst/src ip, dst/src port \| \| pattern_eth_ipv4_gtpu_ipv4_udp \| teid, dst/src ip, dst/src port \| \| pattern_eth_ipv4_gtpu_ipv6 \| teid, dst/src ip \| \| pattern_eth_ipv4_gtpu_ipv6_tcp \| teid, dst/src ip, dst/src port \| \| pattern_eth_ipv4_gtpu_ipv6_udp \| teid, dst/src ip, dst/src port \| \| pattern_eth_ipv6_gtpu_ipv4 \| teid, dst/src ip \| \| pattern_eth_ipv6_gtpu_ipv4_tcp \| teid, dst/src ip, dst/src port \| \| pattern_eth_ipv6_gtpu_ipv4_udp \| teid, dst/src ip, dst/src port \| \| pattern_eth_ipv6_gtpu_ipv6 \| teid, dst/src ip \| \| pattern_eth_ipv6_gtpu_ipv6_tcp \| teid, dst/src ip, dst/src port \| \| pattern_eth_ipv6_gtpu_ipv6_udp \| teid, dst/src ip, dst/src port \| \| pattern_eth_ipv4_gtpu_eh_ipv4 \| teid, qfi, dst/src ip \| \| pattern_eth_ipv4_gtpu_eh_ipv4_tcp\| teid, qfi, dst/src ip, dst/src port\| \| pattern_eth_ipv4_gtpu_eh_ipv4_udp\| teid, qfi, dst/src ip, dst/src port\| \| pattern_eth_ipv4_gtpu_eh_ipv6 \| teid, qfi, dst/src ip \| \| pattern_eth_ipv4_gtpu_eh_ipv6_tcp\| teid, qfi, dst/src ip, dst/src port\| \| pattern_eth_ipv4_gtpu_eh_ipv6_udp\| teid, qfi, dst/src ip, dst/src port\| \| pattern_eth_ipv6_gtpu_eh_ipv4 \| teid, qfi, dst/src ip \| \| pattern_eth_ipv6_gtpu_eh_ipv4_tcp\| teid, qfi, dst/src ip, dst/src port\| \| pattern_eth_ipv6_gtpu_eh_ipv4_udp\| teid, qfi, dst/src ip, dst/src port\| \| pattern_eth_ipv6_gtpu_eh_ipv6 \| teid, qfi, dst/src ip \| \| pattern_eth_ipv6_gtpu_eh_ipv6_tcp\| teid, qfi, dst/src ip, dst/src port\| \| pattern_eth_ipv6_gtpu_eh_ipv6_udp\| teid, qfi, dst/src ip, dst/src port\| +----------------------------------+------------------------------------+ Signed-off-by: Yuying Zhang <yuying.zhang@intel.com> Acked-by: Qi Zhang <qi.z.zhang@intel.com>	2021-04-15 14:22:13 +02:00
Wenzhuo Lu	9c9aa00403	net/iavf: add offload path for Rx AVX512 flex descriptor Add a specific path for RX AVX512 (flexible descriptor). In this path, support the HW offload features, like, checksum, VLAN stripping, RSS hash. This path is chosen automatically according to the configuration. 'inline' is used, then the duplicate code is generated by the compiler. Signed-off-by: Wenzhuo Lu <wenzhuo.lu@intel.com> Acked-by: Qi Zhang <qi.z.zhang@intel.com>	2021-04-14 14:48:06 +02:00
Haifei Luo	bf085dcba1	app/testpmd: add command for single flow dump Add support for single flow dump. The CLIs to dump one rule: flow dump PORT rule ID to dump all: flow dump PORT all Examples: testpmd> flow dump 0 all testpmd> flow dump 0 rule 0 Signed-off-by: Haifei Luo <haifeil@nvidia.com> Acked-by: Ajit Khaparde <ajit.khaparde@broadcom.com>	2021-04-14 13:19:55 +02:00
Haifei Luo	50c383793b	ethdev: dump single flow rule Previous implementations support dump all the flows. Add new arg rte_flow in rte_flow_dev_dump to dump one flow. Signed-off-by: Haifei Luo <haifeil@nvidia.com> Acked-by: Ajit Khaparde <ajit.khaparde@broadcom.com> Acked-by: Ori Kam <orika@nvidia.com>	2021-04-14 13:19:55 +02:00
Li Zhang	74c8ec894d	ethdev: add packet mode in meter profile structure Currently meter algorithms only supports rate is bytes per second (BPS). Add packet_mode flag in meter profile parameters data structure. So that it can meter traffic by packet per second. When packet_mode is 0, the profile rates and bucket sizes are specified in bytes per second and bytes when packet_mode is not 0, the profile rates and bucket sizes are specified in packets and packets per second. The below structure will be extended: rte_mtr_meter_profile rte_mtr_capabilities Signed-off-by: Li Zhang <lizh@nvidia.com> Acked-by: Matan Azrad <matan@nvidia.com> Acked-by: Cristian Dumitrescu <cristian.dumitrescu@intel.com> Acked-by: Jerin Jacob <jerinj@marvell.com> Acked-by: Ajit Khaparde <ajit.khaparde@broadcom.com>	2021-04-13 18:40:58 +02:00
Dong Zhou	cb299214a6	doc: update push/pop VLAN support in mlx5 guide Updates the documentation for push/pop VLAN support. In E-Switch mode, push VLAN on ingress traffic and pop VLAN in egress traffic are both support. Signed-off-by: Dong Zhou <dongzhou@nvidia.com> Reviewed-by: Asaf Penso <asafp@nvidia.com>	2021-04-13 13:37:50 +02:00
Matan Azrad	07b0b75370	cryptodev: formalize key wrap method in API The Key Wrap approach is used by applications in order to protect keys located in untrusted storage or transmitted over untrusted communications networks. The constructions are typically built from standard primitives such as block ciphers and cryptographic hash functions. The Key Wrap method and its parameters are a secret between the keys provider and the device, means that the device is preconfigured for this method using very secured way. The key wrap method may change the key length and layout. Add a description for the cipher transformation key to allow wrapped key to be forwarded by the same API. Add a new feature flag RTE_CRYPTODEV_FF_CIPHER_WRAPPED_KEY to be enabled by PMDs support wrapped key in cipher trasformation. Signed-off-by: Matan Azrad <matan@nvidia.com> Acked-by: Akhil Goyal <gakhil@marvell.com>	2021-04-16 12:43:33 +02:00
Fan Zhang	c21574edc5	cryptodev: add dequeue count parameter in raw API This patch changes the experimental raw data path dequeue burst API. Originally the API enforces the user to provide callback function to get maximum dequeue count. This change gives the user one more option to pass directly the expected dequeue count. Signed-off-by: Fan Zhang <roy.fan.zhang@intel.com> Acked-by: Akhil Goyal <gakhil@marvell.com>	2021-04-16 12:43:33 +02:00
Tejasree Kondoj	398b70cbbb	crypto/octeontx2: support lookaside IPv4 transport mode Adding support for IPv4 lookaside IPsec transport mode. Signed-off-by: Tejasree Kondoj <ktejasree@marvell.com> Acked-by: Akhil Goyal <gakhil@marvell.com>	2021-04-16 12:43:33 +02:00
Tejasree Kondoj	9a1cc8f1ed	examples/ipsec-secgw: support UDP encapsulation Adding lookaside IPsec UDP encapsulation support for NAT traversal. Application has to add udp-encap option to sa config file to enable UDP encapsulation on the SA. Signed-off-by: Tejasree Kondoj <ktejasree@marvell.com> Acked-by: Akhil Goyal <gakhil@marvell.com> Acked-by: Konstantin Ananyev <konstantin.ananyev@intel.com>	2021-04-16 12:43:33 +02:00
Tejasree Kondoj	0ff065d096	crypto/octeontx2: support UDP encapsulation Adding UDP encapsulation support for IPsec in lookaside protocol mode. Signed-off-by: Tejasree Kondoj <ktejasree@marvell.com> Acked-by: Akhil Goyal <gakhil@marvell.com>	2021-04-16 12:43:33 +02:00
Matan Azrad	d014dddb2d	cryptodev: support multiple cipher data-units In cryptography, a block cipher is a deterministic algorithm operating on fixed-length groups of bits, called blocks. A block cipher consists of two paired algorithms, one for encryption and the other for decryption. Both algorithms accept two inputs: an input block of size n bits and a key of size k bits; and both yield an n-bit output block. The decryption algorithm is defined to be the inverse function of the encryption. For AES standard the block size is 16 bytes. For AES in XTS mode, the data to be encrypted\decrypted does not have to be multiple of 16B size, the unit of data is called data-unit. The data-unit size can be any size in range [16B, 2^24B], so, in this case, a data stream is divided into N amount of equal data-units and must be encrypted\decrypted in the same data-unit resolution. For ABI compatibility reason, the size is limited to 64K (16-bit field). The new field dataunit_len is inserted in a struct padding hole, which is only 2 bytes long in 32-bit build. It could be moved and extended later during an ABI-breakage window. The current cryptodev API doesn't allow the user to select a specific data-unit length supported by the devices. In addition, there is no definition how the IV is detected per data-unit when single operation includes more than one data-unit. That causes applications to use single operation per data-unit even though all the data is continuous in memory what reduces datapath performance. Add a new feature flag to support multiple data-unit sizes, called RTE_CRYPTODEV_FF_CIPHER_MULTIPLE_DATA_UNITS. Add a new field in cipher capability, called dataunit_set, where the devices can report the range of the supported data-unit sizes. Add a new cipher transformation field, called dataunit_len, where the user can select the data-unit length for all the operations. All the new fields do not change the size of their structures, by filling some struct padding holes. They are added as exceptions in the ABI check file libabigail.abignore. Using a bitmap to report the supported data-unit sizes capability allows the devices to report a range simply as same as the user to read it simply. also, thus sizes are usually common and probably will be shared among different devices. Signed-off-by: Matan Azrad <matan@nvidia.com> Signed-off-by: Thomas Monjalon <thomas@monjalon.net> Acked-by: Akhil Goyal <gakhil@marvell.com>	2021-04-16 12:43:33 +02:00
Tejasree Kondoj	de5eb0a604	common/cpt: support encrypted digest mode Added support for DIGEST_ENCRYPTED mode for octeontx and octeontx2 platforms. Signed-off-by: Tejasree Kondoj <ktejasree@marvell.com> Acked-by: Akhil Goyal <gakhil@marvell.com>	2021-04-16 12:43:33 +02:00
Stephen Hemminger	9667d97c25	pflock: add phase-fair reader writer locks This is a new type of reader-writer lock that provides better fairness guarantees which better suited for typical DPDK applications. A pflock has two ticket pools, one for readers and one for writers. Phase-fair reader writer locks ensure that neither reader nor writer will be starved. Neither reader or writer are preferred, they execute in alternating phases. All operations of the same type (reader or writer) that acquire the lock are handled in FIFO order. Write operations are exclusive, and multiple read operations can be run together (until a write arrives). A similar implementation is in Concurrency Kit package in FreeBSD. For more information see: "Reader-Writer Synchronization for Shared-Memory Multiprocessor Real-Time Systems", http://www.cs.unc.edu/~anderson/papers/ecrts09b.pdf Signed-off-by: Stephen Hemminger <stephen@networkplumber.org> Acked-by: Honnappa Nagarahalli <honnappa.nagarahalli@arm.com> Acked-by: Konstantin Ananyev <konstantin.ananyev@intel.com>	2021-04-14 21:59:47 +02:00
Pavan Nikhilesh	005e6265e0	doc: announce event Rx adapter config changes The Rx adapter event vector configuration will be merged into Rx adapter queue configuration to simplify enabling event vectorization. Signed-off-by: Pavan Nikhilesh <pbhagavatula@marvell.com> Acked-by: Ray Kinsella <mdr@ashroe.eu> Acked-by: Jerin Jacob <jerinj@marvell.com> Acked-by: Jay Jayatheerthan <jay.jayatheerthan@intel.com>	2021-04-12 09:23:34 +02:00
Pavan Nikhilesh	1cc44d4092	eventdev: introduce event vector capability Introduce rte_event_vector datastructure which is capable of holding multiple uintptr_t of the same flow thereby allowing applications to vectorize their pipeline and reducing the complexity of pipelining the events across multiple stages. This approach also reduces the scheduling overhead on a event device. Add a event vector mempool create handler to create mempools based on the best mempool ops available on a given platform. Signed-off-by: Pavan Nikhilesh <pbhagavatula@marvell.com> Acked-by: Jerin Jacob <jerinj@marvell.com> Acked-by: Ray Kinsella <mdr@ashroe.eu> Acked-by: Jay Jayatheerthan <jay.jayatheerthan@intel.com>	2021-04-12 09:23:34 +02:00
Shijith Thotton	64ea4ae178	event/octeontx2: support timer periodic mode Add support for periodic mode in event timer adapter. Signed-off-by: Shijith Thotton <sthotton@marvell.com> Acked-by: Pavan Nikhilesh <pbhagavatula@marvell.com>	2021-04-12 09:23:34 +02:00
Shijith Thotton	a10d79a60b	eventdev: introduce adapter flags for periodic mode A timer adapter in periodic mode can be used to arm periodic timers. This patch adds flags used to advertise capability and configure timer adapter in periodic mode. Capability flag should be set for adapters which support periodic mode. Below is a programming sequence on the usage: /* check for periodic mode support by reading capability. / rte_event_timer_adapter_caps_get(...); / create adapter in periodic mode by setting periodic flag (RTE_EVENT_TIMER_ADAPTER_F_PERIODIC) and resolution. / rte_event_timer_adapter_create_ext(...); / arm periodic timer of configured resolution / rte_event_timer_arm_burst(...); / timer event will be periodically generated at configured resolution till cancel is called. / while (running) { rte_event_dequeue_burst(...); } / cancel periodic timer which stops generating events */ rte_event_timer_cancel_burst(...); Signed-off-by: Shijith Thotton <sthotton@marvell.com> Acked-by: Erik Gabriel Carrillo <erik.g.carrillo@intel.com> Acked-by: Jerin Jacob <jerinj@marvell.com>	2021-04-12 09:23:34 +02:00
Timothy McDaniel	698fa82941	event/dlb: remove driver Remove event/dlb driver from DPDK code base. Updated release note's removal section to reflect the same. Also updated doc/guides/rel_notes/release_20_11.rst to fix the the missing link issue due to removal of doc/guides/eventdevs/dlb.rst Signed-off-by: Timothy McDaniel <timothy.mcdaniel@intel.com>	2021-04-12 09:21:30 +02:00
Smadar Fuks	1b14508b3b	net/octeontx2: support flow action port ID Action port_id was not supported until now. In this patch the action port_id supports passing from input port PF to output port which is one of input port respective VF Signed-off-by: Smadar Fuks <smadarf@marvell.com> Acked-by: Jerin Jacob <jerinj@marvell.com>	2021-04-08 17:09:57 +02:00
Salem Sol	fd44e8288f	net/mlx5: support NVGRE encap action in sampling Add support for NVGRE encap as a sample action and validate it. Signed-off-by: Salem Sol <salems@nvidia.com> Acked-by: Viacheslav Ovsiienko <viacheslavo@nvidia.com>	2021-04-08 01:09:24 +02:00
Salem Sol	be47c9819f	net/mlx5: support VXLAN encap action in sampling Add support for VXLAN encap as a sample action and validate it. Signed-off-by: Salem Sol <salems@nvidia.com> Acked-by: Viacheslav Ovsiienko <viacheslavo@nvidia.com>	2021-04-08 01:09:24 +02:00
Pallavi Kadam	ce6617a069	net/ice: build on Windows - Add Intel ice PMD support on Windows. - Remove #include sys/ioctl header file as it is not needed. - Replace x86intrin.h with rte_vect.h to avoid __m_prefetchw conflicting types. - Replace POSIX usleep() API with rte API. - Add a new macro for the access() API as the original function has been deprecated on Windows. - Add extra cflags '-fno-asynchronous-unwind-tables' to avoid MinGW build error: Error: invalid register for .seh_savexmm - Add documentation to support ice PMD on Windows. Update the release notes and features list for the same. Signed-off-by: Pallavi Kadam <pallavi.kadam@intel.com> Reviewed-by: Ranjit Menon <ranjit.menon@intel.com> Acked-by: Jie Zhou <jizh@microsoft.com> Reviewed-by: Ferruh Yigit <ferruh.yigit@intel.com>	2021-04-06 19:00:36 +02:00
Min Hu (Connor)	38b539d96e	net/hns3: support IEEE 1588 PTP Add hns3 support for new ethdev APIs to enable and read IEEE1588/ 802.1AS PTP timestamps. Signed-off-by: Min Hu (Connor) <humin29@huawei.com>	2021-04-01 18:39:55 +02:00
Ashwin Sekhar T K	2da3159197	mempool/cnxk: add build infra and doc Add the meson based build infrastructure for Marvell CNXK mempool driver along with stub implementations for mempool device probe. Also add Marvell CNXK mempool base documentation. Signed-off-by: Pavan Nikhilesh <pbhagavatula@marvell.com> Signed-off-by: Jerin Jacob <jerinj@marvell.com> Signed-off-by: Nithin Dabilpuram <ndabilpuram@marvell.com> Signed-off-by: Ashwin Sekhar T K <asekhar@marvell.com>	2021-04-09 08:32:24 +02:00
Nithin Dabilpuram	68a03efeed	doc: add Marvell cnxk platform guide Platform specific guide for Marvell OCTEON CN9K/CN10K SoC is added. Signed-off-by: Nithin Dabilpuram <ndabilpuram@marvell.com> Signed-off-by: Jerin Jacob <jerinj@marvell.com>	2021-04-09 08:32:24 +02:00
Suanming Mou	330a70b773	regex/mlx5: add data path scattered mbuf process UMR (User-Mode Memory Registration) WQE can present data buffers scattered within multiple mbufs with single indirect mkey. Take advantage of the UMR WQE, scattered mbuf in one operation can be presented to an indirect mkey. The RegEx which only accepts one mkey can now process the whole scattered mbuf in one operation. The maximum scattered mbuf can be supported in one UMR WQE is now defined as 64. The mbufs from multiple operations can be combined into one UMR WQE as well if there is enough space in the KLM array, since the operations can address their own mbuf's content by the mkey's address and length. However, one operation's scattered mbuf's can't be placed in two different UMR WQE's KLM array, if the UMR WQE's KLM does not has enough free space for one operation, the extra UMR WQE will be engaged. In case the UMR WQE's indirect mkey will be over wrapped by the SQ's WQE move, the mkey's index used by the UMR WQE should be the index of last the RegEX WQE in the operations. As one operation consumes one WQE set, build the RegEx WQE by reverse helps address the mkey more efficiently. Once the operations in one burst consumes multiple mkeys, when the mkey KLM array is full, the reverse WQE set index will always be the last of the new mkey's for the new UMR WQE. In GGA mode, the SQ WQE's memory layout becomes UMR/NOP and RegEx WQE by interleave. The UMR and RegEx WQE can be called as WQE set. The SQ's pi and ci will also be increased as WQE set not as WQE. For operations don't have scattered mbuf, uses the mbuf's mkey directly, the WQE set combination is NOP + RegEx. For operations have scattered mbuf but share the UMR WQE with others, the WQE set combination is NOP + RegEx. For operations complete the UMR WQE, the WQE set combination is UMR + RegEx. Signed-off-by: John Hurley <jhurley@nvidia.com> Signed-off-by: Suanming Mou <suanmingm@nvidia.com> Acked-by: Ori Kam <orika@nvidia.com>	2021-04-08 22:52:55 +02:00
Thomas Monjalon	4d509afa7b	pci: rename catch-all ID The name of the constant PCI_ANY_ID was missing RTE_ prefix. It is renamed, and the old name becomes a deprecated alias. While renaming, the duplicate definitions in rte_bus_pci.h are removed to keep only those in rte_pci.h. Note: rte_pci.h is included in rte_bus_pci.h Signed-off-by: Thomas Monjalon <thomas@monjalon.net> Reviewed-by: Parav Pandit <parav@nvidia.com> Reviewed-by: David Marchand <david.marchand@redhat.com>	2021-04-06 14:52:49 +02:00
Junfeng Guo	f91c7b2da1	net/iavf: support GTPU inner IPv4 for flow director Support GTPU_(EH)_IPV4 inner L3 and L4 fields matching for AVF FDIR. +------------------------------+---------------------------------+ \| Pattern \| Input Set \| +------------------------------+---------------------------------+ \| eth/ipv4/gtpu/ipv4 \| inner: src/dst ip \| \| eth/ipv4/gtpu/ipv4/udp \| inner: src/dst ip, src/dst port \| \| eth/ipv4/gtpu/ipv4/tcp \| inner: src/dst ip, src/dst port \| \| eth/ipv4/gtpu/eh/ipv4 \| inner: src/dst ip \| \| eth/ipv4/gtpu/eh/ipv4/udp \| inner: src/dst ip, src/dst port \| \| eth/ipv4/gtpu/eh/ipv4/tcp \| inner: src/dst ip, src/dst port \| \| eth/ipv4/gtpu/eh(0)/ipv4 \| inner: src/dst ip \| \| eth/ipv4/gtpu/eh(0)/ipv4/udp \| inner: src/dst ip, src/dst port \| \| eth/ipv4/gtpu/eh(0)/ipv4/tcp \| inner: src/dst ip, src/dst port \| \| eth/ipv4/gtpu/eh(1)/ipv4 \| inner: src/dst ip \| \| eth/ipv4/gtpu/eh(1)/ipv4/udp \| inner: src/dst ip, src/dst port \| \| eth/ipv4/gtpu/eh(1)/ipv4/tcp \| inner: src/dst ip, src/dst port \| +------------------------------+---------------------------------+ Signed-off-by: Junfeng Guo <junfeng.guo@intel.com> Acked-by: Qi Zhang <qi.z.zhang@intel.com>	2021-03-30 01:21:17 +02:00
Jiawen Wu	f611dada1a	net/txgbe: update link setup process of backplane NICs Add device arguments to support runtime options. And use these configuration to control the link setup flow, to adapt to different NIC's construction. Use firmware version to control the impact of firmware update. And fix some left bugs. Signed-off-by: Jiawen Wu <jiawenwu@trustnetic.com>	2021-03-29 17:49:34 +02:00
Dmitry Kozlyuk	8d96d605d3	net/vmxnet3: enable on Windows Remove OS restriction and update release notes. For the record, tested on the following setup with Windows Server 2019 in QEMU (-device vmxnet3) : [ping ] [ ] [ ping] [OS ] [ dpdk-skeleton ] [ OS] [virtio---]--sockets--[---vmxnet3 vmxnet3---]--sockets--[---virtio] [Debian VM] [ Windows VM ] [Debian VM] Debian VMs successfully ping'd each other with Windows forwarding. Signed-off-by: Dmitry Kozlyuk <dmitry.kozliuk@gmail.com> Acked-by: Yong Wang <yongwang@vmware.com>	2021-03-25 17:00:49 +01:00
Hongbo Zheng	63e05f19b8	net/hns3: support Rx descriptor status query Add support for query Rx descriptor status in hns3 driver. Check the descriptor specified and provide the status information of the corresponding descriptor. Signed-off-by: Hongbo Zheng <zhenghongbo3@huawei.com> Signed-off-by: Min Hu (Connor) <humin29@huawei.com>	2021-03-23 13:04:33 +01:00
Hongbo Zheng	656a6d9cc0	net/hns3: support Tx descriptor status query Add support for query Tx descriptor status in hns3 driver. Check the descriptor specified and provide the status information of the corresponding descriptor. Signed-off-by: Hongbo Zheng <zhenghongbo3@huawei.com> Signed-off-by: Min Hu (Connor) <humin29@huawei.com>	2021-03-23 13:04:33 +01:00
Chengchang Tang	d0ab89e633	net/hns3: support outer UDP checksum Kunpeng930 support outer UDP cksum, this patch add support for it. Signed-off-by: Chengchang Tang <tangchengchang@huawei.com> Signed-off-by: Min Hu (Connor) <humin29@huawei.com>	2021-03-23 13:04:32 +01:00
Chengwen Feng	a124f9e959	net/hns3: add runtime config to select IO burst function Currently, the driver support multiple IO burst function and auto selection of the most appropriate function based on offload configuration. Most applications such as l2fwd/l3fwd don't provide the means to change offload configuration, so it will use the auto selection's io burst function. This patch support runtime config to select io burst function, which add two config: rx_func_hint and tx_func_hint, both could assign vec/sve/simple/common. The driver will use the following rules to select io burst func: a. if hint equal vec and meet the vec Rx/Tx usage condition then use the neon function. b. if hint equal sve and meet the sve Rx/Tx usage condition then use the sve function. c. if hint equal simple and meet the simple Rx/Tx usage condition then use the simple function. d. if hint equal common then use the common function. e. if hint not set then: e.1. if meet the vec Rx/Tx usage condition then use the neon function. e.2. if meet the simple Rx/Tx usage condition then use the simple function. e.3. else use the common function. Note: the sve Rx/Tx usage condition based on the vec Rx/Tx usage condition and runtime environment (which must support SVE). In the previous versions, driver will preferred use the sve function when meet the sve Rx/Tx usage condition, but in this case driver could get better performance if use the neon function. Signed-off-by: Chengwen Feng <fengchengwen@huawei.com> Signed-off-by: Min Hu (Connor) <humin29@huawei.com>	2021-03-23 13:04:32 +01:00
Ed Czeck	6c7f491e7f	net/ark: generalize meta data between FPGA and PMD In this commit we generalize the movement of user-specified meta data between mbufs and FPGA AXIS tuser fields using user-defined hook functions. - Previous use of PMD dynfields are removed - Remove emptied rte_pmd_ark.h - Hook function added to ark_user_ext - Add hook function calls in Rx and Tx paths - Update guide with example of hook function use Signed-off-by: Ed Czeck <ed.czeck@atomicrules.com>	2021-03-22 16:56:27 +01:00
Ed Czeck	f2764c3688	net/ark: cleanup dynamic extension interface - Rename extension functions with rte_pmd_ark prefix - Update local function documentation Signed-off-by: Ed Czeck <ed.czeck@atomicrules.com>	2021-03-22 16:56:27 +01:00
Ed Czeck	9ee9e0d3b8	net/ark: update to reflect FPGA updates - New PCIe IDs using net/ark driver - Update Version IDs and structures specified by hardware - New internal descriptor status for TX - Adjust data placement in RX operations, headroom in retained for segmented mbufs Signed-off-by: Ed Czeck <ed.czeck@atomicrules.com>	2021-03-22 16:56:27 +01:00
Tal Shnaiderman	1325a1ffd9	eal: rename thread TLS API Rename the key opaque pointer from rte_tls_key to rte_thread_key to avoid confusion with transport layer security. Also rename and remove the "_tls" term from the following functions to avoid redundancy: rte_thread_tls_key_create rte_thread_tls_key_delete rte_thread_tls_value_set rte_thread_tls_value_get Suggested-by: Vladimir Medvedkin <vladimir.medvedkin@intel.com> Suggested-by: Morten Brørup <mb@smartsharesystems.com> Signed-off-by: Tal Shnaiderman <talshn@nvidia.com> Acked-by: Morten Brørup <mb@smartsharesystems.com>	2021-03-26 09:22:39 +01:00
Bruce Richardson	e34e2f55a7	telemetry: make the legacy registration function internal The function for registration of callbacks for legacy telemetry was documented as internal-only in the API documents, but marked as experimental in the version.map file. Since this is an internal-only function, for consistency we update the version mapping to have it as internal. Signed-off-by: Bruce Richardson <bruce.richardson@intel.com> Acked-by: Ciara Power <ciara.power@intel.com>	2021-03-25 17:35:10 +01:00

1 2 3 4 5 ...

1900 Commits