WebRTC-based Smart City Management Operation and Maintenance Methods and Systems
By using WebRTC-based multi-terminal audio and video instant messaging services, the problem of limited number of maintenance personnel and expert guidance in smart city management has been solved, enabling real-time information sharing and security management, and improving maintenance efficiency.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2022-03-31
- Publication Date
- 2026-03-17
AI Technical Summary
The operation and maintenance management in smart city management is inefficient. The number of operation and maintenance personnel is limited and their capabilities vary. Expert guidance is restricted, and on-site personnel cannot obtain remote technical support in a timely manner, which affects the efficiency of operation and maintenance management.
It adopts a WebRTC-based multi-terminal audio and video instant messaging service to achieve real-time audio and video communication through terminals and servers, create multi-user rooms, provide audio and video conferencing functions, and add identity tags and content tags to users for recording, storage and security classification management of audio and video information.
It improves the efficiency of smart city management and operation, enables real-time audio and video communication among multiple users, facilitates information sharing between experts and maintenance personnel, reduces system resource consumption, and ensures information security.
Smart Images

Figure CN114666527B_ABST
Abstract
Description
Technical Field
[0001] This invention relates to the field of video call technology, specifically to a smart city management operation and maintenance method and system based on WebRTC. Background Technology
[0002] Smart city management is a new urban management model supported by next-generation information technology and in the context of knowledge society innovation 2.0. It achieves comprehensive and thorough perception, ubiquitous broadband interconnection, and intelligent integrated application through next-generation information technology support, and promotes people-oriented sustainable innovation characterized by user innovation, open innovation, mass innovation, and collaborative innovation. Smart city management is an important component of smart cities.
[0003] Smart city management utilizes sensing technology to monitor and comprehensively perceive all aspects of urban management. It employs various sensing devices and intelligent monitoring systems for intelligent identification, providing a three-dimensional understanding of changes in the urban environment and location. The system integrates, processes, and analyzes the data collected by these devices, enabling intelligent integration with business processes and facilitating proactive responses from the monitoring center. Sensing devices are deployed throughout the city, including in areas such as road lighting, landscape lighting, transformer substations, manhole covers, underground pipe networks, hazardous sources, low-lying areas with water accumulation, bridges, tunnels, slopes, roadside trees, ancient and famous trees, parks, reservoirs, and garbage transfer stations. This allows for real-time monitoring of urban infrastructure. The more sensing devices deployed and the wider their coverage, the better the real-time monitoring effect and the more conducive it is to the operation of smart city management. However, the sensing devices themselves also require maintenance. More sensing devices mean more maintenance personnel are needed, and the urban municipal components monitored by the sensing devices also require maintenance personnel. However, the number of maintenance personnel is limited and their capabilities vary, which seriously restricts the efficiency of maintenance management in smart city management. In particular, when encountering difficult maintenance problems, ordinary maintenance personnel cannot solve them and require guidance from more capable experts. However, experts are often limited by the number and location of personnel and cannot conduct on-site diagnosis. On-site maintenance personnel cannot obtain remote technical support in a timely manner. In addition, command and dispatch personnel are also limited by the number and location of personnel and cannot conduct on-site command. All of these factors greatly affect the efficiency of smart city management maintenance. Summary of the Invention
[0004] One of the objectives of this invention is to provide a WebRTC-based smart city management operation and maintenance method, which provides multi-terminal audio and video instant communication services for smart city management operation and maintenance, thereby improving the work efficiency of smart city management operation and maintenance.
[0005] The basic solution provided by this invention is a smart city management and operation method based on WebRTC, which includes the following:
[0006] Obtain user information to log in;
[0007] Send a connection establishment request;
[0008] It receives and responds to connection establishment requests, upgrades the communication protocol from HTTP to WebSocket for real-time communication, gathers multiple users engaged in real-time communication, and creates rooms.
[0009] Obtain audio and video information;
[0010] The system processes audio and video information and transmits it to the corresponding users, enabling audio and video communication conferences between users in the room.
[0011] The beneficial effects of Basic Solution 1: In smart city management, users can use terminals to conduct real-time communication based on WebRTC through a server. WebRTC can be implemented in a webpage, is low-cost, and aggregates multiple terminals for real-time communication, creating rooms where users can conduct audio and video conferencing. This method can be applied to smart city management operation and maintenance programs, providing multi-terminal audio and video instant communication services for on-site maintenance personnel, monitoring hall command personnel, and remote diagnostic guidance experts. It enables multi-terminal users to conduct real-time audio and video communication, providing intuitive and visual real-time audio and video information exchange services for monitoring hall command personnel, on-site maintenance personnel, and remote diagnostic guidance experts anytime, anywhere, thereby improving the efficiency of smart city management operation and maintenance.
[0012] Furthermore, the method also includes:
[0013] Add identity tags to users; these tags include: commanders, maintenance personnel, and guidance experts, and a user can have multiple identity tags.
[0014] Detect the identity tags of users speaking in the current room;
[0015] Determine whether the speaker's identity tag is "guidance expert". If so, record audio and video information for both the guidance expert and the operations and maintenance personnel.
[0016] The audio and video information are integrated into a single audio-video file and stored according to a preset storage format.
[0017] Manage content tags for stored audio and video information; content tag management includes: adding content tags, deleting content tags, and modifying content tags;
[0018] When terminals make audio and video calls in a room, they identify whether the audio information contains preset keywords, where the preset keywords are content tags. If so, the corresponding audio and video information is retrieved and pushed.
[0019] It receives audio and video information selection signals and plays the audio and video information.
[0020] Beneficial effects: Adding identity tags to users facilitates user identification. The system detects the identity tags of users speaking in the current room and determines if the user's tag information corresponds to a guidance expert. If so, audio and video information is recorded for both the guidance expert and maintenance personnel. This records the guidance expert's audio and video information for later use, while avoiding recording audio and video information from other irrelevant users, reducing memory usage and saving system resources. During audio and video calls between terminals in a room, the system identifies whether the audio contains preset keywords, which are content tags. If so, the corresponding audio and video information is retrieved and pushed. The system also receives audio and video information selection signals and plays the audio and video information. This eliminates the need for users to manually search for relevant audio and video information during meetings, allowing for timely retrieval. Both guidance experts and maintenance personnel can use historical audio and video information as reference and guidance, reducing the workload of guidance experts. Maintenance personnel can receive guidance based on the audio and video information and directly and clearly see how to operate. Guidance experts can also adjust their guidance based on the audio and video, achieving better guidance results and improving maintenance effectiveness and efficiency.
[0021] Furthermore, the method also includes:
[0022] Enter meeting information;
[0023] After detecting the identity tags of each user in the room within a preset time period, determine whether there are any users whose identity tags are "guidance experts". If not, retrieve and push audio and video information related to the meeting information based on the meeting information.
[0024] It receives audio and video information selection signals and plays the audio and video information.
[0025] Beneficial effects: The preset time period is set according to the meeting information. If the expert cannot start the guidance on time, and the expert has not entered the room by the preset time period, the relevant audio and video information can be retrieved and pushed to the maintenance personnel. This may allow the maintenance personnel to learn the relevant content in advance, or perform maintenance based on the audio and video information.
[0026] Furthermore, the method also includes:
[0027] Set user permission values;
[0028] Set audio and video information permission values;
[0029] After receiving the audio / video information selection signal, the method further includes:
[0030] Determine if the user's permission value is greater than or equal to the audio / video information permission value. If yes, play the audio / video information; otherwise, prompt the user that they do not have permission to view it.
[0031] Beneficial effects: This allows for security grading of audio and video information. Users with access levels lower than the required level cannot view audio and video information. For audio and video information involving core technologies or confidential information, it is not allowed to be viewed by any maintenance personnel, thus preventing the leakage of core technologies or confidential information.
[0032] The second objective of this invention is to provide a WebRTC-based smart city management and maintenance management system, which provides multi-terminal audio and video instant communication services for smart city management and maintenance, thereby improving the work efficiency of smart city management and maintenance.
[0033] This invention provides a second basic solution: a smart city management and maintenance management system based on WebRTC, comprising: a server and a terminal;
[0034] The terminal is used to obtain user information for user login and send connection establishment requests to the server; it is also used to obtain audio and video information and transmit it to the server.
[0035] The server is used to receive and respond to connection requests, upgrade the communication protocol from HTTP to WebSocket, enable real-time communication between the server and terminals, aggregate multiple terminals engaged in real-time communication, and create rooms. It is also used to process audio and video information and transmit it to the corresponding terminals, enabling audio and video communication conferencing between terminals within the room.
[0036] The beneficial effects of Basic Solution Two: In smart city management, users can use terminals to conduct real-time communication based on WebRTC through the service. WebRTC can be implemented in a webpage, is low-cost, and users only need to use a terminal to conduct real-time communication. The server will aggregate multiple terminals conducting real-time communication and create rooms, allowing users to conduct audio and video communication conferences through the rooms. This system can be integrated into the smart city management operation and maintenance program, providing multi-terminal audio and video instant communication services for on-site operation and maintenance personnel, command personnel in the monitoring hall, and remote diagnostic guidance experts. It realizes the audio and video instant communication function for multiple users, providing intuitive and visual real-time audio and video information exchange services for command personnel in the monitoring hall, on-site operation and maintenance personnel, and remote diagnostic guidance experts at any time and in any area, thereby improving the efficiency of smart city management operation and maintenance.
[0037] Furthermore, the server includes: an audio codec module, a video codec module, and a network transmission module;
[0038] The audio codec module is used to process and encode the audio information transmitted by the terminal, generate audio data, and send it to the corresponding terminal through the network transmission module; it is also used to decode the received audio data, obtain audio information, and send it to the corresponding terminal through the network transmission module.
[0039] The video encoding / decoding module is used to process and encode the video information transmitted by the terminal, generate video data, and send it to the corresponding terminal through the network transmission module; it is also used to decode the video data received by the network transmission module, obtain video information, and send it to the corresponding terminal.
[0040] The network transmission module is used to transmit audio and video data to the terminal.
[0041] Beneficial effects: The audio codec module, video codec module, and network transmission module work together to process and transmit audio and video information. Furthermore, the audio and video codec modules ensure that audio decoding is prioritized when network conditions are poor, preventing both audio and video transmission from being interrupted.
[0042] Furthermore, the server also includes: an identity tag management module, a detection module, a judgment module, and a storage module;
[0043] The identity tag management module is used to add identity tags to users; the identity tags include: command personnel, operation and maintenance personnel, and guidance experts, and a user can add multiple identity tags;
[0044] The detection module is used to detect the identity tags of users speaking in the current room;
[0045] The judgment module is used to determine whether the identity tag information of the speaking user is a guidance expert. If so, the recording module of the speaking user's terminal and the recording module of the operation and maintenance personnel's terminal are started.
[0046] The storage module is used to integrate audio and video information into audio-video information for storage according to a preset storage format;
[0047] The terminal includes: a recording module;
[0048] The recording module is used to record audio and video information and send it to the storage module for storage.
[0049] Beneficial effects: Adding identity tags to users facilitates user identification. The system detects the identity tags of users speaking in the current room and determines whether the speaking user's identity tag is that of a guidance expert. If so, the recording modules of the speaking user's terminal and the maintenance personnel's terminal are activated to record the guidance expert's audio and video information for later use. This avoids recording the audio and video information of other irrelevant users, reducing memory usage and saving system resources.
[0050] Furthermore, the server also includes: a content tag management module and a recognition module;
[0051] The content tag management module is used to manage the content tags of audio and video information stored in the storage module; the management includes: adding content tags, deleting content tags, and modifying content tags;
[0052] The recognition module is used to identify whether the audio information contains preset keywords when terminals make audio and video calls in a room. The preset keywords are content tags. If so, the corresponding audio and video information is retrieved and pushed.
[0053] The terminal also includes: a playback module;
[0054] The playback module is used to receive audio and video information selection signals and play audio and video information.
[0055] Beneficial effects: When terminals conduct audio and video calls in a room, the recognition module identifies whether the audio information contains preset keywords, which are content tags. If so, it retrieves the corresponding audio and video information and pushes it. The playback module receives the audio and video information selection signal and plays the audio and video information. Therefore, users do not need to search for relevant audio and video information during the meeting; they can retrieve relevant audio and video information in a timely manner. Both guidance experts and maintenance personnel can use historical audio and video information as a reference and guidance, reducing the workload of guidance experts. Maintenance personnel can receive guidance based on the audio and video information and can also directly and clearly see how to operate. Guidance experts can also make adjustments based on the audio and video information, thereby achieving better guidance effects, improving maintenance effectiveness, and increasing maintenance efficiency.
[0056] Furthermore, the terminal also includes: an input module;
[0057] The input module is used to input meeting information;
[0058] The detection module is also used to detect the identity tags of each user in the room after a preset time period;
[0059] The judgment module is also used to determine whether there is a user whose identity tag is "guidance expert". If not, it retrieves audio and video information related to the meeting information and pushes it to the terminal of the user whose identity tag is "operation and maintenance personnel".
[0060] Beneficial effects: The preset time period is set according to the meeting information. In cases where the expert cannot start the guidance on time, if the expert has not entered the room by the preset time period, the relevant audio and video information can be retrieved and pushed to the maintenance personnel. This may allow the maintenance personnel to learn the relevant content in advance, or to perform maintenance based on the audio and video information.
[0061] Furthermore, the server also includes: a permission management module;
[0062] The permissions management module is used to set user permission values;
[0063] The content tag management module is also used to set audio and video information permission values;
[0064] The playback module is also used to determine whether the user's permission value is greater than or equal to the audio and video information permission value. If so, the audio and video information is played; if not, the user is prompted that they do not have permission to view it.
[0065] Beneficial effects: This allows for security grading of audio and video information. Users with access levels lower than the required level cannot view audio and video information. For audio and video information involving core technologies or confidential information, it is not allowed to be viewed by any maintenance personnel, thus preventing the leakage of core technologies or confidential information. Attached Figure Description
[0066] Figure 1 This is a flowchart illustrating an embodiment of the WebRTC-based smart city management operation and maintenance method of the present invention.
[0067] Figure 2 This is a logical block diagram of an embodiment of the WebRTC-based smart city management and maintenance system of the present invention. Detailed Implementation
[0068] The following detailed description illustrates the specific implementation method:
[0069] Example 1
[0070] The basic implementation examples are as follows: Figure 1 As shown: The smart city management operation and maintenance method based on WebRTC includes the following:
[0071] Obtain user information to log in;
[0072] Send a connection establishment request;
[0073] It receives and responds to connection establishment requests, upgrades the communication protocol from HTTP to WebSocket for real-time communication, gathers multiple users engaged in real-time communication, and creates a room. Users within a room are interconnected, and a terminal in a room can select other terminals, i.e., other users, to join the room.
[0074] Obtain audio and video information;
[0075] The system processes audio and video information and transmits it to the corresponding users, enabling audio and video communication conferences between users in the room; the processing includes encoding and decoding.
[0076] Add identity tags to users; these tags include: commanders, maintenance personnel, and guidance experts, and a user can have multiple identity tags.
[0077] Set user permission values;
[0078] Detect the identity tags of users speaking in the current room;
[0079] Determine whether the speaker's identity tag is "guidance expert". If so, record audio and video information for both the guidance expert and the operations and maintenance personnel.
[0080] The audio and video information are integrated into a single audio-video file and stored according to a preset storage format.
[0081] Set audio and video information permission values;
[0082] Manage content tags for stored audio and video information; content tag management includes: adding content tags, deleting content tags, and modifying content tags;
[0083] When terminals make audio and video calls in a room, they identify whether the audio information contains preset keywords, where the preset keywords are content tags. If so, the corresponding audio and video information is retrieved and pushed.
[0084] The system receives audio / video information selection signals and plays the audio / video information; or it receives audio / video information selection signals and determines whether the user's user permission value is greater than or equal to the audio / video information permission value. If so, the audio / video information is played; if not, the user is prompted that they do not have permission to view it. This allows for security classification of audio / video information. Users with permission values lower than the audio / video information permission value cannot view the audio / video information. For some audio / video information involving core technologies or secrets, it is not allowed to be viewed by any maintenance personnel to prevent the leakage of core technologies or secrets.
[0085] In addition, this method also includes: inputting meeting information; wherein the meeting information includes: meeting topic, meeting category, and meeting summary; the meeting information can be input before sending the connection establishment request;
[0086] After detecting the identity tags of each user in the room within a preset time period, determine whether there are any users whose identity tags are "guidance experts". If not, retrieve and push audio and video information related to the meeting information based on the meeting information.
[0087] The system receives an audio / video information selection signal and plays the audio / video information; or it receives an audio / video information selection signal, determines whether the user's permission value is greater than or equal to the audio / video information permission value, and if so, plays the audio / video information; otherwise, it prompts the user that they do not have permission to view it.
[0088] In smart city management, users can use terminals to conduct real-time communication based on WebRTC through a server. WebRTC can be implemented in a webpage, is low-cost, and users only need a terminal to conduct real-time communication. The server aggregates multiple terminals conducting real-time communication and creates rooms, allowing users to conduct audio and video conferencing between rooms. This method can be applied to smart city management operation and maintenance programs to provide multi-terminal audio and video instant communication services for on-site operation and maintenance personnel, command personnel in the monitoring hall, and remote diagnostic guidance experts. It enables multi-terminal users to conduct real-time audio and video communication, providing intuitive and visual real-time audio and video information exchange services for command personnel in the monitoring hall, on-site operation and maintenance personnel, and remote diagnostic guidance experts anytime and in any area, thereby improving the efficiency of smart city management operation and maintenance.
[0089] User identity tags are set to clearly identify users in each audio and video communication meeting. Different identity tags can be reset for different meetings. Based on the identity tags, audio and screen recordings are made when the expert speaks, and the maintenance personnel's terminals are also recorded. This records the expert's guidance content and the maintenance personnel's maintenance process under the expert's guidance, completing a full maintenance process record. Audio and video information is generated and stored. Users can view the audio and video information in the storage module through their terminals, and can also use the audio and video information and guidance documents as teaching materials, which is beneficial for the training of maintenance personnel and experts.
[0090] When users conduct audio and video communication, the system identifies whether the audio contains preset keywords, which are content tags. If so, the corresponding audio and video information is retrieved and pushed to the terminals of guidance experts and maintenance personnel. Guidance experts and maintenance personnel can select audio and video information. The playback module receives the audio and video information selection signal and plays the audio and video information. Thus, guidance experts and maintenance personnel can use historical audio and video information as a reference and guidance, reducing the workload of guidance experts. Maintenance personnel can receive guidance based on the audio and video information and can also directly and clearly see how to operate. Guidance experts can also make adjustments based on the audio and video information, thereby achieving better guidance effects and improving maintenance effectiveness and efficiency.
[0091] Example 2
[0092] This embodiment is basically as shown in the appendix. Figure 2 As shown: A smart city management and maintenance management system based on WebRTC, including: server and terminal;
[0093] The terminal is used to obtain user information for user login and send a connection establishment request to the server; it is also used to obtain audio and video information and transmit them to the server; in this embodiment, the terminal includes: mobile phone, computer and tablet;
[0094] The terminal includes: a recording module, a playback module, and an input module;
[0095] The recording module is used to record audio and video information and send it to the storage module for storage;
[0096] The playback module is used to receive audio and video information selection signals and play audio and video information; it is also used to determine whether the user's user permission value is greater than or equal to the audio and video information permission value. If so, the audio and video information is played; if not, the user is prompted that they do not have permission to view it. This allows for security classification of audio and video information. Users with user permission values lower than the audio and video information permission value cannot view audio and video information. For some audio and video information involving core technologies or secrets, it is not allowed to be viewed by any maintenance personnel to prevent the leakage of core technologies or secrets.
[0097] The input module is used to input meeting information, which includes: meeting topic, meeting category, and meeting summary. The input module can be used to input meeting information before the terminal sends a connection request to the server.
[0098] The server receives and responds to connection requests, upgrades the communication protocol from HTTP to WebSocket, enables real-time communication between the server and terminals, aggregates multiple terminals engaged in real-time communication, and creates rooms. Terminals within a room can select other terminals (i.e., other users) to join the room, and all terminals within a room are interconnected. The server also processes audio and video information and transmits it to the corresponding terminals, facilitating audio and video conferencing between terminals within the room. In this embodiment, the server can be a collection of various servers, such as a WebSocket server, a room server, and a signaling server.
[0099] The server includes: an audio encoding / decoding module, a video encoding / decoding module, a network transmission module, an identity tag management module, a detection module, a judgment module, a storage module, a content tag management module, a recognition module, a permission management module, a subtitle generation module, and a guidance document generation module;
[0100] The audio codec module is used to process and encode the audio information transmitted by the terminal, generate audio data, and send it to the corresponding terminal through the network transmission module; it is also used to decode the received audio data, obtain audio information, and send it to the corresponding terminal through the network transmission module.
[0101] The video encoding / decoding module is used to process and encode the video information transmitted by the terminal, generate video data, and send it to the corresponding terminal through the network transmission module; it is also used to decode the video data received by the network transmission module, obtain video information, and send it to the corresponding terminal; in this embodiment, the video encoding / decoding module adopts VP8 encoding, so that high-quality video can be achieved with less data communication;
[0102] The network transmission module is used to transmit audio and video data to the terminal. In this embodiment, the network transmission module integrates the RTP / SPRT protocol stack and the STUN / TURN / ICE protocol. The RTP / SPRT protocol stack is the transmission protocol, and the STUN / TURN / ICE protocol is used to solve NAT traversal.
[0103] The identity tag management module is used to add identity tags to users to facilitate user identification; the identity tags include: command personnel, operation and maintenance personnel, and guidance experts, and a user can add multiple identity tags;
[0104] The detection module is used to detect the identity tags of users speaking in the current room;
[0105] The judgment module is used to determine whether the identity tag information of the speaking user is a guidance expert. If so, the recording module of the speaking user's terminal and the recording module of the operation and maintenance personnel's terminal are started.
[0106] The storage module is used to integrate audio and video information into audio-video information for storage according to a preset storage format. In this embodiment, the preset storage format limits the file name of the audio-video information, including the recording time and meeting information. This records the guidance audio-video information of the expert, which is convenient for subsequent use, and does not record the audio-video information of other irrelevant users, thereby reducing memory usage and saving system resources.
[0107] The content tag management module is used to manage the content tags of audio and video information stored in the storage module. The management includes adding, deleting and modifying content tags. It is also used to set audio and video information permission values. The content tags are input through the terminal's input module.
[0108] The recognition module is used to identify whether the audio contains preset keywords when terminals make audio and video calls in a room. The preset keywords are content tags. If so, the corresponding audio and video information is retrieved and pushed. The corresponding audio and video information refers to audio and video information whose content tags contain the preset keywords.
[0109] The detection module is also used to detect the identity tags of each user in the room after a preset time period;
[0110] The judgment module is also used to determine whether there is a user with the identity tag of "guidance expert". If not, it retrieves audio and video information related to the meeting information and pushes it to the terminal of the user with the identity tag of "operations and maintenance personnel". The preset time period is set according to the meeting information. In the case of guidance experts not being able to start guidance on time, if the guidance expert has not entered the room by the preset time period, the relevant audio and video information can be retrieved and pushed to the operations and maintenance personnel first. This may allow the operations and maintenance personnel to learn the relevant content in advance, or allow the operations and maintenance personnel to perform operations and maintenance based on the audio and video information. Specifically, retrieving audio and video information related to the meeting information involves: extracting the same information as the content tags of all audio and video information in the meeting information; and then retrieving audio and video information whose content tags contain the same information.
[0111] The permissions management module is used to set user permission values;
[0112] The subtitle generation module is used to automatically generate subtitles from the audio and video information stored in the storage module;
[0113] The instruction document generation module is used to extract subtitles and generate instruction documents.
[0114] In smart city management, users can use terminals to conduct real-time communication based on WebRTC through a server. WebRTC can be implemented in a webpage, is low-cost, and users only need a terminal to conduct real-time communication. The server aggregates multiple terminals conducting real-time communication and creates rooms, allowing users to conduct audio and video conferencing between rooms. This system can be integrated into smart city management operation and maintenance programs, providing multi-terminal audio and video instant communication services for on-site operation and maintenance personnel, command personnel in the monitoring hall, and remote diagnostic guidance experts. It enables multi-terminal users to conduct real-time audio and video communication, providing intuitive and visual real-time audio and video information exchange services for command personnel in the monitoring hall, on-site operation and maintenance personnel, and remote diagnostic guidance experts anytime, anywhere, thereby improving the efficiency of smart city management operation and maintenance.
[0115] This system assigns user identity tags to clearly identify users in each audio and video communication conference. Different identity tags can be reset for different conferences. Based on the identity tags, audio and screen recordings are made when the expert speaks, and the maintenance personnel's terminals also record the recordings. This records the expert's guidance content and the maintenance personnel's maintenance process under the expert's guidance, completing a full maintenance process record. The system generates and stores the audio and video information, which users can view through the terminal's playback module. The audio and video information and guidance documents can also be used as teaching materials, which is beneficial for training maintenance personnel and experts.
[0116] When users conduct audio and video communication, the recognition module identifies whether the audio contains preset keywords, which are content tags. If so, the corresponding audio and video information is retrieved and pushed to the terminals of guidance experts and maintenance personnel. Guidance experts and maintenance personnel can select audio and video information. The playback module receives the audio and video information selection signal and plays the audio and video information. Thus, guidance experts and maintenance personnel can use historical audio and video information as a reference and guidance, reducing the workload of guidance experts. Maintenance personnel can receive guidance based on the audio and video information and can also directly and clearly see how to operate. Guidance experts can also make adjustments based on the audio and video, thereby achieving better guidance effects, improving maintenance effectiveness, and increasing maintenance efficiency.
[0117] The above descriptions are merely embodiments of the present invention. Commonly known structures and characteristics are not described in detail here. Those skilled in the art are aware of all common technical knowledge in the field prior to the application date or priority date, are aware of all existing technologies in that field, and have the ability to apply conventional experimental methods prior to that date. Those skilled in the art can, under the guidance of this application, improve and implement this solution in combination with their own capabilities. Some typical known structures or methods should not be obstacles for those skilled in the art to implement this application. It should be noted that those skilled in the art can make several modifications and improvements without departing from the structure of the present invention. These should also be considered within the scope of protection of the present invention, and will not affect the effectiveness of the implementation of the present invention or the practicality of the patent. The scope of protection claimed in this application should be determined by the content of its claims, and the specific embodiments described in the specification can be used to interpret the content of the claims.
Claims
1. A WebRTC-based intelligent urban management operation and maintenance management method, comprising: Acquire user information to log in; Send a connection request; It is characterized in that: Receive and respond to the connection request, upgrade the communication protocol from HTTP to WebSocket, conduct real-time communication, and gather multiple users for real-time communication and create a room; Acquire audio and video information; Process the audio and video information and transmit them to the corresponding user, and conduct audio and video communication between users in the room; Add an identity tag to the user; the identity tag includes a commander, an operator, and a guide expert; Detect the identity tag of the current speaker in the room; Determine whether the identity tag of the speaker is a guide expert; if so, record the audio and video information of the guide expert and the operator; Integrate the audio and video information into audio and video information according to a preset storage format and store them; Manage the content tags of the stored audio and video information; the content tag management includes adding, deleting, and modifying content tags; When the terminals conduct audio and video communication in the room, identify whether the audio information contains a preset keyword; the preset keyword is a content tag; if so, retrieve the corresponding audio and video information and push it; Receive the audio and video information selection signal and play the audio and video information; It also includes: Input meeting information; Detect the identity tags of the users in the room after a preset time period, determine whether there is a user with an identity tag of a guide expert, and if not, retrieve and push the audio and video information related to the meeting information according to the meeting information; Receive the audio and video information selection signal and play the audio and video information; It also includes: Set a user permission value; Set an audio and video information permission value for the audio and video information; After receiving the audio and video information selection signal, it also includes: Determine whether the user permission value is greater than or equal to the audio and video information permission value; if so, play the audio and video information; if not, prompt the user that they do not have permission to view.
2. The smart urban management operation and maintenance management system based on WebRTC, comprising: Server and terminal; The terminal acquires user information to log in and sends a connection request to the server; It is also used to acquire audio and video information and transmit them to the server; The server receives and responds to the connection request, upgrades the communication protocol from HTTP to WebSocket, conducts real-time communication with the terminal, and gathers multiple terminals for real-time communication and creates a room; it is also used to process the audio and video information and transmit them to the corresponding terminal, and conduct audio and video communication between terminals in the room; The server also includes an identity tag management module, a detection module, a judgment module, and a storage module; the terminal includes a recording module; The identity tag management module adds an identity tag to the user; the identity tag includes a commander, an operator, and a guide expert; The detection module detects the identity tag of the current speaker in the room; The judgment module determines whether the identity tag of the speaker is a guide expert; if so, it starts the recording module of the terminal of the speaker and the recording module of the terminal of the operator; The storage module is configured to store audio information and video information in a preset storage format as audio-video information; The terminal further comprises an input module; The input module is configured to input conference information; The detection module is further configured to detect the identity tags of the users in the room after a preset time period; The judgment module is further configured to determine whether there is a user with an identity tag of a guide expert, and if not, to retrieve audio-video information related to the conference information and push the audio-video information to the terminal of the user with an identity tag of an operation and maintenance personnel according to the conference information; The server further comprises a permission management module and a content tag management module; The permission management module is configured to set a user permission value; The content tag management module is further configured to set an audio-video information permission value for the audio-video information; The terminal further comprises a playing module; The playing module is further configured to determine whether the user permission value of the user is greater than or equal to the audio-video information permission value, and if so, to play the audio-video information, and if not, to prompt the user that the user has no permission to view. 3.The WebRTC-based intelligent urban management operation and maintenance management system according to claim 2, characterized in that: The server comprises an audio codec module, a video codec module and a network transmission module; The audio codec module is configured to process and encode the audio information transmitted by the terminal to generate audio data, and transmit the audio data to the corresponding terminal through the network transmission module, and further configured to decode the received audio data to obtain the audio information, and transmit the audio information to the corresponding terminal through the network transmission module; The video codec module is configured to process and encode the video information transmitted by the terminal to generate video data, and transmit the video data to the corresponding terminal through the network transmission module, and further configured to decode the video data received by the network transmission module to obtain the video information, and transmit the video information to the corresponding terminal; The network transmission module is configured to transmit the audio data and the video data to the terminal. 4.The WebRTC-based intelligent urban management operation and maintenance management system according to claim 3, characterized in that: The server further comprises an identification module; The content tag management module is configured to manage the content tags of the audio-video information stored in the storage module, including adding, deleting and modifying the content tags; The identification module is configured to identify whether the preset keyword is contained in the audio information when the terminals perform audio-video communication in the room, the preset keyword being a content tag, and if so, to retrieve the corresponding audio-video information and push the audio-video information; The playing module is configured to receive an audio-video information selection signal and play the audio-video information.
Citation Information
Patent Citations
A video acquisition management method and a server
CN109005382A
Nuclear power station remote monitoring, diagnosis, maintenance integration system
CN204733292U