Method for optimizing the failure detection of redundancy protocols using test data packets

DE502017016914D1Active Publication Date: 2025-07-03HIRSCHMANN AUTOMATION AND CONTROL GMBH
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
DE502017016914
Authority / Receiving Office
DE · DE
Patent Type
Patents
Current Assignee / Owner
Priority Date
2016-12-16
Filing Date
2017-12-15
Publication Date
2025-07-03
Estimated Expiration
2037-12-15

AI Technical Summary

Technical Problem

In network operations, the transmission of test data packets for failure detection can be significantly delayed due to competing with payload data, leading to increased worst-case detection and failover times in the event of a network failure.

Method used

The method combines dynamic redundancy protocols with test data packets and frame preemption or time slotting techniques to reduce the worst-case detection time of network failures. This is achieved by prioritizing the transmission of test data packets, either by interrupting payload data transmission or reserving dedicated time slots for test data packets.

Benefits of technology

The approach significantly reduces the worst-case detection and failover times by minimizing the dwell time of test data packets in network devices, ensuring quick failure detection and switchover to redundant paths.

✦ Generated by Eureka AI based on patent content.
Patent Text Reader
Need to check novelty before this filing date? Find Prior Art

Description

[0001] The invention relates to a method for operating a network, wherein network devices in the network exchange user data with each other via at least one transmission medium by transmitting user data packets and at least one redundancy protocol is used to reduce the risk of failure, wherein this at least one redundancy protocol carries out a transmission of test data packets to detect failures in the network, according to the features of the respective preamble of the two independent patent claims.

[0002] Methods for operating a network are known, wherein network devices in the network exchange data with each other and from and to other devices such as sensors, actuators, and the like via at least one transmission medium, and at least one redundancy protocol is used to reduce the risk of failure. This at least one redundancy protocol carries out a cyclical transmission of test data packets to detect failures in the network. Networks, in particular Ethernet data networks, form the technological basis, for example, for industrial monitoring and for control networks in production lines. A failure or malfunction in these networks is typically associated with a loss of productivity or control or monitoring reliability.

[0003] Further prior art is disclosed in US 2009 / 161562 A1, EP 1 734 700 A1 and EP 3 451 591 A1. 1

[0004] To reduce the risk of failure, redundancy protocols such as the Media Redundancy Protocol (MRP) or the Device Level Ring (DLR) have proven effective. These redundancy protocols typically require the exchange of payload data (also referred to as productive data) between network devices and other devices in the network using payload data packets (also referred to as frames), as well as the cyclic transmission of test data packets (also referred to as test frames or test packets) to detect network failures.However, depending on the amount of data that the network devices exchange with each other and possibly also with the other devices in the network, and the associated utilization of at least one transmission medium due to the transmitted payload data packets, the transmission of test data packets can be significantly delayed, which negatively affects the worst-case switching time to redundant network paths in the event of a failure.

[0005] The use of frame preemption according to IEEE802.3br and IEEE 802.1Qbu also enables the interruption of the transmission of other data (payload data) in order to improve the latency for certain traffic classes of the payload data packets.

[0006] According to IEEE 802.1Qbv, it is known to use a time-division multiple access (TDMA) method to define communication cycles and access to the transmission medium based on class-of-service priorities (CoS) in a virtual local area network (VLAN) header of Ethernet frames in order to implement hard real-time requirements with low latencies and deviations (jitter).

[0007] The test data packets of a redundancy protocol typically compete with other data (payload) for access to at least one transmission medium (e.g., data line, radio link, or the like; in Ethernet applications, primarily wired). As the available bandwidth usage increases and the number of network devices in the network increases, particularly in a ring network, the payload can increasingly lead to delays in the forwarding of test data packets. The worst-case detection and switchover times of redundancy protocols with test data packets are therefore determined by the maximum delay in forwarding a test frame on each network device.

[0008] The invention is therefore based on the object of providing a method for operating a network that avoids the disadvantages described above. In particular, the time required to switch to another transmission path when an interruption of at least one transmission medium is detected is to be reduced.

[0009] This problem is solved by the features of the independent patent claims.

[0010] The present invention combines known methods of dynamic redundancy protocols with test data packets and the use of frame preemption or time slotting techniques in three alternative or combinable ways. This, either individually or in combination, enables a significant reduction in the worst-case detection time of a network failure and thus a reduction in the worst-case failover time in the event of a fault.

[0011] To reduce the risk of failure, the redundancy protocols MRP and DLR, for example, have proven effective in ring networks. Here, a network device configured as a ring master monitors the network by regularly sending test data packets through the ring. The test data packets are received by every network device participating in the redundancy protocol and forwarded until they arrive back at the ring master. If the test data packets fail to arrive, a failure occurs in the ring network, and an alternative transmission path is activated.

[0012] According to a first solution, the invention provides that the transmission of a user data packet is interrupted and, instead of further transmission of this user data packet, a test data packet is transmitted and only then is the transmission of the remaining user data packet carried out.

[0013] By using frame preemption, redundancy protocols can treat test data packets as express data, and the transmission of payload packets can be interrupted. This interruption enables the prioritized transmission of test data packets, even if a payload packet has already begun. After the test data packet has been transmitted, the transmission of payload packets is resumed and completed.

[0014] By interrupting the payload data transmission and prioritizing the forwarding of the test data packets, the dwell time of the test data packets in the network devices participating in the redundancy protocol is reduced to a minimum defined by the frame preemption mechanism. This results in the dwell time of the test data packets being independent of the payload data length, thus significantly reducing worst-case detection and failover times in the event of a fault.

[0015] According to a first alternative of a second solution, the invention provides that a predeterminable time range is reserved for the transmission of test data packets, wherein no payload data packets (regardless of the traffic class or the priority assigned to them) are transmitted within the predeterminable time range. Thus, a predefined time range is always available for the transmission of a test data packet, which is at least large enough for the test data packet to be transmitted within the predefined time range. It is also possible to select a longer time for the reserved transmission of test data packets than is required for the actual transmission. In this case, it is ensured that the test data packets are transmitted with the highest priority and precedence over the transmission of payload data packets.However, if the specified time range is longer than the time required to transmit a test data packet, the available bandwidth will not be used optimally, since within the specified time range (time slot) after the transmission of a test data packet there is still time available to transmit payload data packets, in particular payload data packets of lower priority.

[0016] According to a second alternative of the second solution, the invention therefore provides that a predeterminable time period is reserved for the transmission of test data packets, wherein the transmission of payload data can continue within the predeterminable time period and the test data packets are transmitted with a higher transmission priority than the payload data packets. In such a case, if time is still available within this time slot after the priority transmission of a test data packet within the predefined time period, payload data packets, in particular payload data packets with a lower priority or payload data packets that are in a queue, can also be transmitted within this time slot. This means that the bandwidth is optimally utilized and it is not necessary to wait for the predefined time period to expire before further payload data packets can be transmitted.

[0017] In one embodiment, it is provided that a test data packet is first transmitted by the network devices and only then does at least one network device begin transmitting a user data packet.

[0018] The fact that no user data packets are transmitted within the predefined time range (also referred to as a time window or time slot) advantageously ensures that the user data packets cannot cause any influence or delay for the transmission of the test data packets.

[0019] When using time slotting techniques such as IEEE802.1Qbv, the delay in forwarding a test data packet (test frame) on a network device due to a payload data packet (payload frame) already in transit can even be completely eliminated. Test data packets are assigned to a class, a so-called traffic class, based on recognizable properties of these packets. A predefined time slot is then configured for this class using time slotting techniques such as IEEE802.1Qbv. In such a time slot, the test data packets are then transmitted with priority, either within a class and / or across classes. The advantage of this solution lies in the optimization of worst-case detection and switching times and thus in the use of time slotting techniques such as the aforementioned "Enhancements for Scheduled Traffic" (IEEE802.1Qbv) to allow the transmission of test data packets on the ring using dedicated time slots without additional waiting time in the forwarding network device. A predefined portion of the available network bandwidth is thus reserved for test frames (test data packets). For optimal use of the failure detection period, the ring master sends the test frames synchronized with the start of the time slot designated for test frames.

[0020] In a further development of the invention, the generation of the test data packets in the network device configured as the master is time-coupled with the opening of the time slot designated for the test data packets. This ensures optimal use of the time window for transmitting the test data packet.

[0021] Due to the length independence, in contrast to the known state of the art, the use of jumbo frames (i.e. extra-long data packets) is also possible without affecting the worst-case detection and switching times.

[0022] Finally, the invention is not limited to ring redundancy protocols, but extends to redundancy protocols based on the use of test data packets, regardless of the network topology.

[0023] According to the invention, the detection of a failure (interruption of transmission for whatever reason, such as a cable break, pulled or defective connector, power failure in a network device or the like) is thus always advantageously quickly detected by minimizing the waiting time of test data packets on network devices, so that the time until a failure is detected is minimized.

[0024] An example calculation can illustrate the contribution of the invention. In DIN EN 62439-2:2010-09 (MRP), Chapter 9.5.4, a worst-case calculation is performed for 50 ring participants, resulting in a switching time of 26.2 ms. If the value for the smallest framelet with frame preemption of 64 bytes / 5.12 µs is used for T Queue, the resulting time is only 14.5 ms. By using the time slot method, T Queue can be reduced to 0 µs, resulting in a switching time of only 14.0 ms.

[0025] The two solutions described in general terms above are described in more detail below with reference to the figures using an exemplary embodiment.

[0026] Figure 1shows an example of a network in the form of a ring topology in which four network devices (NWG) are present. These network devices (NWG) are connected to each other via a wired transmission medium, in particular a data line (DL), for the purpose of transmitting data. It is assumed that in order to apply a redundancy protocol, such as the Media Redundancy Protocol (MRP) or the Device Level Ring (DLR), one network device acts as the master (in Figure 1 designated as M), while the remaining network devices NWG are configured as clients (in the Figure 1(referred to as R1, R2, and R3). In a conventional manner, if an error is detected within the network using the transmitted test data packets, other transmission sections can be switched over to the transmission of the payload data (and also the test data packets), so that, based on the known redundancy protocols, another network device (NWG) that was previously configured as a client can also assume the function of a master. These switching mechanisms, which are also used here, are generally known, so there is no need to go into them in more detail here.

[0027] In the Figure 1 Four network devices NWG are shown as an example, although in practice there are often more than four network devices NWG, rarely fewer than four network devices NWG.

[0028] In Figure 2 is the passage of test data packets TP within the ring topology according to Figure 1The network device NWG, which is configured as master M, sends a test data packet TP to the next network device NWG (here the client R1). From there, the client R1 sends the test data packet to the next network device NWG, namely the next client R2. Once the test data packet TP has been received here, it is forwarded to the next network device NWG, namely the client R3, which can forward the received test data packet to the master M. Due to this mechanism, it is detected in a conventional manner that the ring is closed and there is no interruption.

[0029] It should be noted that this approach is described using a ring topology. However, the invention is not limited to ring topologies and can be applied to other network topologies, such as line topologies.

[0030] In Figure 2The ideal case of the passage of test data packets through the ring network is shown, which does not yet take into account the exchange of payload data in the form of payload data packets. This is therefore a theoretical ideal case that will not occur in practice, as it does not consider the transmission of payload data packets within the network.

[0031] In Figure 3 The case is taken into account that not only the test data packets are transmitted via the network devices, but also that user data packets are transmitted between the individual network devices and across them.

[0032] In the Figure 3The worst case scenario for this transmission of test data packets and payload data packets is that the master M sends a test data packet. Since the client R1 processes, and in particular sends, a payload data packet NP1, the test data packet TP received by the master M can only be forwarded once the payload data packet NP1 has been completely transmitted. The same applies to the other network devices R2 and R3, so that in any case, the forwarding of the test data packet by the other network devices NWG (here the clients R2 and R3) is delayed due to the processing or transmission of the other payload data packets NP2 and NP3.

[0033] According to the first solution of the invention, as described in Figure 4shown, the transmission of a payload packet is interrupted and instead of further transmission of this payload packet, a test data packet is transmitted and only then the transmission of the remaining payload packet is carried out. With regard to the Figure 4 This means that the first network device (the master M, which does not necessarily have to be the first network device, but can be any other network device) sends out a test data packet and the next network device NWG, here the client R1, starts sending a payload packet 1. According to the invention, however, the system does not wait until the payload packet NP1 has been completely transmitted by the network device R1, but rather interrupts the transmission of this payload packet NP1 in order to initiate the transmission of the test data packet TP by the network device R1. After this has happened, the remaining payload packet NP1 (in the Figure 4the larger part of NP1) is emitted.

[0034] The same process occurs on network device R2, which has already begun transmitting a payload packet NP2 upon receiving the test data packet. Once the test data packet TP from network device R1 has been received by network device R2, the transmission of the payload packet NP2 that had already begun is interrupted, and the test data packet TP is transmitted by network device R2. After this has occurred, the remaining part of the payload packet NP2 (here, again, the larger part, for example) is transmitted further.

[0035] This continues with the further network device R3, so that the first solution approach according to the invention clearly shows the significant reduction in the transmission time of a test data packet TP on the ring network compared to the worst case, which in Figure 3 is shown.

[0036] The length or size of the respective payload data packet NP1, NP2 or NP3 according to Figure 4 depends on the time at which the respective test data packet TP was received on the respective network device. This means that the length or size of the payload packets NP1, NP2 and NP3 before and after the test data packet TP can be the same or different than shown in Figure 4 is shown.

[0037] In Figure 5 An alternative to the second solution according to the invention is shown using an embodiment in which a predeterminable time range (time slot TP) is reserved for all network devices and a test data packet is first transmitted by the network devices and only then at least one network device begins to transmit a user data packet. This means with regard to the Figure 5that a time slot is reserved for the test data packets (Slot TP), so that a test data packet is always transmitted over the network before any payload data packets are transmitted.

[0038] In the example case according to Figure 5The master M therefore sends out its test data packet TP within the reserved time slot, which is received and forwarded by the client R1. The same applies to the test data packet TP sent by the client R1 and received by the client R2, as well as to the test data packet TP sent by the client R2 and received by the client R3. Only when the test data packets TP have been transmitted within this reserved time slot can at least one other network device, in this example the client R1, send out its payload packet NP1. The same also applies to the two other network devices R2 and R3, which can only send their payload packets NP2 and NP3 respectively when the test data packet TP has been transmitted within the reserved time slot (slot TP).During the time period reserved by the time slot, no payload data packets can be transmitted, even if the test data packet has already been transmitted. The payload data packets must therefore wait until their assigned time window opens.

[0039] In Figure 6 The general case of the second solution according to the invention is shown. In this case, a predeterminable time range is reserved for the transmission of test data packets, with no payload data packets being transmitted within the predeterminable time range. This can thus be at least two or more time slots (as opposed to one time slot for all network devices) and even different time slots on different network devices to account for path and processing latencies.

[0040] In Figure 6It is thus shown that each network device (M, R1 to R3) is assigned a reserved time slot (slot TP) in which the test data packet TP can be transmitted. The transmission of the payload packets is possible at any time before and after this reserved time slot. Thus, in this embodiment, it is shown that the network device configured as master M sends a test data packet to the client R1 in a time slot reserved for this purpose. Before and after the reserved time slot, the network device configured as master can receive and send payload packets. The client R1, in turn, has reserved a time slot within which it can forward the received test data packet TP. Figure 6It is shown that the payload packet NP1 from client R1 is sent after the reserved time slot for the test data packet TP. The same applies to client R2. Client R3 has also reserved a time slot for the test data packet TP. However, this client R3 can, in turn, send its payload packet NP3 before the reserved time slot.

[0041] The time slots reserved by the respective network devices are identical (i.e., they have the same length). Alternatively, different time slots can be reserved for the test data packets for each network device or group of network devices (configured in the network device). It is important to ensure that the specified time range (time slot) has a minimum length sufficient for the secure and complete transmission of a test data packet.

[0042] In this alternative of the second solution according to the invention, the test data packets are transmitted preferentially in the reserved time slot of the network devices, so that either the transmission of a network data packet must occur before the reserved time slot or only after the transmission of the test data packet within its reserved time slot. This advantageously ensures that whenever a test data packet is pending transmission, no payload data packet is in transmission and would interfere with the transmission of the test data packet.

[0043] This second solution also results in a significantly faster transmission of the test data packets (especially when compared with the Figure 3 ), so that in the event of an error, the system can react and switch over much more quickly.

[0044] In the second solution according to the invention (in one of the two or both alternatives), the test data packets are treated as so-called express data, just as in the first solution according to the invention, whereby in the first solution the payload data packets are interrupted and in the second solution the test data packets have the highest priority and thus "free" travel on the network.

Claims

1. Method for operating a network, wherein network devices in the network exchange useful data with one another via at least one transmission medium through the transmission of useful data packets and at least one redundancy protocol is applied in order to reduce a failure risk, wherein this at least one redundancy protocol performs a transmission of test data packets in order to detect failures in the network, characterized in that the transmission of a useful data packet is interrupted and, instead of the further transmission of this useful data packet, a test data packet is transmitted in order to detect failures in the network and only thereafter is the transmission of the remaining useful data packet carried out.

2. Method for operating a network, wherein network devices in the network exchange useful data with one another via at least one transmission medium through the transmission of useful data packets and at least one redundancy protocol is applied in order to reduce a failure risk, wherein this at least one redundancy protocol performs a transmission of test data packets in order to detect failures in the network, characterized in that a predefinable time slot is reserved for the transmission of test data packets in order to detect failures in the network, wherein no useful data packets are transmitted within the predefinable time slot.

3. Method for operating a network, wherein network devices in the network exchange useful data with one another via at least one transmission medium through the transmission of useful data packets and at least one redundancy protocol is applied in order to reduce a failure risk, wherein this at least one redundancy protocol performs a transmission of test data packets in order to detect failures in the network, characterized in that a predefinable time slot is reserved for the transmission of test data packets in order to detect failures in the network, wherein the transmission of useful data can continue to take place within the predefinable time slot and the test data packets are transmitted with a higher transmission priority compared to the useful data packets.

4. Method for operating a network according to Claim 2 or 3, characterized in that a test data packet is first transmitted by the network device which is configured as the master and only thereafter does at least one network device start to transmit a useful data packet.

5. Method for operating a network according to Claim 2 or 3, characterized in that generation of the test data packets in the network device which is configured as the master is temporally linked to the opening of the time slot provided for test data packets.

6. Method for operating a network according to one of the preceding claims, characterized in that test data packets for detecting failures in the network are given a higher priority, and waiting times for test data packets on network devices are thus reduced to a minimum which is necessary until a lower-priority useful data packet has been interrupted.