WO2024256023A1 - Resource block scheduling in wireless communication network - Google Patents
Resource block scheduling in wireless communication network Download PDFInfo
- Publication number
- WO2024256023A1 WO2024256023A1 PCT/EP2023/066263 EP2023066263W WO2024256023A1 WO 2024256023 A1 WO2024256023 A1 WO 2024256023A1 EP 2023066263 W EP2023066263 W EP 2023066263W WO 2024256023 A1 WO2024256023 A1 WO 2024256023A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- rbg
- ues
- rbs
- rbgs
- target
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Classifications
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04B—TRANSMISSION
- H04B7/00—Radio transmission systems, i.e. using radiation field
- H04B7/02—Diversity systems; Multi-antenna system, i.e. transmission or reception using multiple antennas
- H04B7/04—Diversity systems; Multi-antenna system, i.e. transmission or reception using multiple antennas using two or more spaced independent antennas
- H04B7/0413—MIMO systems
- H04B7/0452—Multi-user MIMO systems
Definitions
- the present disclosure relates generally to the field of wireless communications.
- the present disclosure relates to a Resource Block (RB) scheduling apparatus and method in a wireless communication network, as well as to a corresponding computer program product.
- RB Resource Block
- a properly configured machine learning (ML) model can learn to efficiently allocate Resource Blocks (RBs) (i.e., time-frequency resources) or Resource Block Groups (RBGs) (each typically including 4-16 continuous RBs) to User Equipments (UEs).
- RBs Resource Blocks
- RBGs Resource Block Groups
- UEs User Equipments
- a ML-based RB scheduler can replace the existing RB schedulers relying on heuristic algorithms. More specifically, it has been also shown that the ML-based RB scheduler can outperform the heuristic RB schedulers in spectral efficiency.
- the ML-based RB scheduler has been proved to dominate in scheduling execution time, by doing real-time RB scheduling decisions for uplink approximately 70-80% faster than the existing RB schedulers using pretty complex state-of-the-art heuristic uplink scheduling algorithms.
- an RB scheduling apparatus in a wireless communication network comprises at least one processor and at least one memory storing instructions that, when executed by the at least one processor, cause the apparatus to operate at least as follows.
- the apparatus obtains a list of UEs for which a set of RBs or RBGs is to be scheduled for a target Transmission Time Interval (TTI).
- TTI Transmission Time Interval
- the set of RBs or RBGs is shared by a set of MU-MIMO layers supported by a network node in the wireless communication network.
- Each MU-MIMO layer of the set of MU-MIMO layers is assigned to a different UE of the list of UEs.
- the apparatus obtains a data vector for each RB or RBG of the set of RBs or RBGs for each MU-MIMO layer of the set of MU-MIMO layers.
- the data vector comprises a frequency-specific value and a spatial correlation value.
- the frequency-specific value indicates a channel state of each UE of the list of UEs.
- the spatial correlation value indicates whether the UEs of the list of UEs are spatially separable by using beamforming.
- the apparatus schedules the set of RBs for the list of UEs for the target TTI by using a ML model.
- the ML model comprises at least one Neural Network (NN) configured to receive the data vectors as input data and output a scheduling decision sequentially for each RB or RBG of the set of RBs or RBGs.
- the scheduling decision indicates whether the RB or RBG is to be scheduled for one or more UEs of the list of UEs.
- the apparatus thus configured may take MU-MIMO aspects (i.e., the spatial diversity of the UEs) into account when scheduling (both downlink (DL) and uplink (UL)) RBs or RBGs for the UEs.
- the apparatus may perform the frequency and spatial domain RB scheduling regardless of: a maximum number of UEs to be scheduled to RBs or RBGs, a physical layer numerology/bandwidth (i.e., a number of RBs or RBGs per TTI), (digital/analog/hybrid) conventional and/or ML-based beamformers used in the MU -Ml MO network, and a number of Radio Frequency (RF) chains determining a maximum number of spatially scheduled UEs.
- a maximum number of UEs to be scheduled to RBs or RBGs a physical layer numerology/bandwidth (i.e., a number of RBs or RBGs per TTI), (digital/analog/hybrid) conventional and/or ML-based beamformers used in the MU -Ml MO network
- RF Radio Frequency
- the spatial correlation value comprises at least one of: a singular value decomposition of combined precoders of the UEs of the list of UEs, a Power Normalization Loss (PLN) value calculated for each UE of the list of UEs, and an estimation of a throughput increase caused by scheduling each UE to an additional MU-MIMO layer.
- PPN Power Normalization Loss
- the NN(s) may obtain efficient (in terms of spatial domain (SD) scheduling) decisions for the set of RBs or RBGs.
- the frequency-specific value comprises at least one of: Channel State Information (CSI) in the RB or RBG, a past average DL throughput in the RB or RBG, a past average UL throughput in the RB or RBG, a number of already scheduled other RBs or RBGs of the set of RBs or RBGs, an average size of transmitted DL packets in the RB or RBG, an average size of transmitted UL packets in the RB or RBG, a Modulation and Coding Scheme (MCS), and a scheduling decision made for the previous RB or RBG of the set of RBs or RBGs.
- MCS Modulation and Coding Scheme
- the CSI comprises at least one of: a Radio Resource Management (RRM) measurement for the RB or RBG, a RRM measurement for a sub-band comprising the RB or RBG, and a RRM measurement averaged over an entire bandwidth comprising the sub-band.
- the RRM measurement may comprise a Signal-to- Interference-plus-Noise (SINR), a Received Signal Received Power (RSRP), a Received Signal Received Quality (RSRQ), a Channel Quality Indicator (CQI), and/or any other signal quality parameter.
- SINR Signal-to- Interference-plus-Noise
- RSRP Received Signal Received Power
- RSRQ Received Signal Received Quality
- CQI Channel Quality Indicator
- the at least one NN is trained based on a sequence of tuples by using a reinforcement learning algorithm.
- Each tuple of the sequence of tuples comprises: a data vector obtained for a target RB or RBG of the set of RBs or RBGs for a target MU-MIMO layer of the set of MU-MIMO layers for the target TTI; a scheduling decision outputted by the at least one NN for the target RB or RBG; a performance reward resulted from the scheduling decision for the target RB or RBG; and a data vector obtained for one of: (i) a subsequent RB or RBG of the set of RBs or RBGs for the target MU -Ml MO layer for the target TTI; (ii) the target RB or RBG for a subsequent MU-MIMO layer of the set of MU-MIMO layers for the target TTI; and (iii) the target RB or RBG for the target MU-MIMO layer for a subsequent
- the fourth element in each tuple can be either a next RB for the same MU- MIMO layer, the same RB for a next MU-MIMO layer, or the same RB for a next TTI, it is possible to train the NN(s) to achieve a high long-term reward by performing seamless transitions over either the RBs or RBGs, or the MU-MIMO layers, or the TTIs (in all cases, the training performance is similar, but if the transitions are performed over the RBs or RBGs, it may provide better performance when training continuous RB or RBG allocations, for example, for single carrier waveforms, such as DFT-s-OFDM).
- the at least one NN comprises a set of NNs each configured to output the scheduling decision for each RB or RBG of the set of RBs or RBGs but for a different MU-MIMO layer of the set of MU-MIMO layers.
- the at least one NN comprises a single NN configured to output the scheduling decision for each RB or RBG of the set of RBs or RBGs for each MU-MIMO layer of the set of MU-MIMO layers.
- This embodiment can be used when it is impossible to use a different NN for each MU-MIMO layer (e.g., due to computing resource exhaustion).
- an RB scheduling method in a wireless communication network starts with the step of obtaining a list of UEs for which a set of RBs or RBGs is to be scheduled for a target TTI.
- the set of RBs or RBGs is shared by a set of MU-MIMO layers supported by a network node in the wireless communication network.
- Each MU-MIMO layer of the set of MU-MIMO layers is assigned to a different UE of the list of UEs.
- the method proceeds to the step of obtaining a data vector for each RB or RBG of the set of RBs or RBGs for each MU-MIMO layer of the set of MU-MIMO layers.
- the data vector comprises frequency-specific information and a spatial correlation value.
- the frequency-specific information indicates a channel state of each UE of the list of UEs.
- the spatial correlation value indicates whether the UEs of the list of UEs are spatially separable by using beamforming.
- the method goes on to the step of scheduling the set of RBs or RBGs for the list of UEs for the target TTI by using an ML model.
- the ML model comprises at least one NN configured to receive the data vectors as input data and output a scheduling decision sequentially for each RB or RBG of the set of RBs or RBGs.
- the scheduling decision indicates whether the RB or RBG is to be scheduled for one or more UEs of the list of UEs.
- MU-MIMO aspects i.e., the spatial diversity of the UEs
- the frequency and spatial domain RB or RBG scheduling may be performed regardless of: a maximum number of UEs to be scheduled to RBs or RBGs, a physical layer numerology/bandwidth (i.e., a number of RBs or RBGs per TTI), (digital/analog/hybrid) conventional and/or ML-based beamformers used in the MU-MIMO network, and a number of RF chains determining a maximum number of spatially scheduled UEs.
- the spatial correlation value comprises at least one of: a singular value decomposition of combined precoders of the UEs of the list of UEs; and a PLN value calculated for each UE of the list of UEs.
- the NN(s) may obtain efficient (in terms of the SD scheduling) decisions for the set of RBs or RBGs.
- the frequency-specific information comprises at least one of: CSI in the RB or RBG, a past average DLthroughput in the RB or RBG, a past average UL throughput in the RB or RBG, a number of already scheduled other RBs or RBGs of the set of RBs or RBGs, an average size of transmitted DL packets in the RB or RBG, an average size of transmitted UL packets in the RB or RBG, a MCS, and a scheduling decision made for the previous RB or RBG of the set of RBs or RBGs.
- the NN(s) may obtain the most efficient (in terms of the combined FD and SD scheduling) decisions for the set of RBs or RBGs.
- the CSI comprises at least one of: a RRM measurement for the RB or RBG, a RRM measurement for a sub-band comprising the RB or RBG, and a RRM measurement averaged over an entire bandwidth comprising the sub-band.
- the RRM measurement may comprise a SINR, a RSRP, a RSRQ, a CQI, and/or any other signal quality parameter.
- the at least one NN is trained based on a sequence of tuples by using a reinforcement learning algorithm.
- Each tuple of the sequence of tuples comprises: a data vector obtained for a target RB or RBG of the set of RBs or RBGs for a target MU -Ml MO layer of the set of MU-MIMO layers for the target TTI; a scheduling decision outputted by the at least one NN for the target RB or RBG; a performance reward resulted from the scheduling decision for the target RB or RBG; and a data vector obtained for one of: (i) a subsequent RB or RBG of the set of RBs or RBGs for the target MU-MIMO layer for the target TTI, (ii) the target RB or RBG for a subsequent MU-MIMO layer of the set of MU- MIMO layers for the target TTI, and (iii) the target RB or RBG for the target MU-MIMO layer for a
- the fourth element in each tuple can be either a next RB for the same MU-MIMO layer, the same RB for a next MU-MIMO layer, or the same RB for a next TTI, it is possible to train the NN(s) to achieve a high long-term reward by performing seamless transitions over either the RBs or RBGs, or the MU-MIMO layers, or the TTIs (in all cases, the training performance is similar, but if the transitions are performed over the RBs or RBGs, it may provide better performance when training continuous RB or RBG allocations, for example, for single carrier waveforms, such as DFT-s-OFDM).
- the at least one NN comprises a set of NNs each configured to output the scheduling decision for each RB or RBG of the set of RBs or RBGs but for a different MU-MIMO layer of the set of MU-MIMO layers.
- the at least one NN comprises a single NN configured to output the scheduling decision for each RB or RBG of the set of RBs or RBGs for each MU-MIMO layer of the set of MU-MIMO layers.
- This embodiment can be used when it is impossible to use a different NN for each MU-MIMO layer (e.g., due to computing resource exhaustion).
- a computer program product comprises a computer-readable storage medium that stores a computer code. Being executed by at least one processor, the computer code causes the at least one processor to perform the method according to the second aspect.
- an RB scheduling apparatus in a wireless communication network comprises a means for obtaining a list of UEs for which a set of RBs or RBGs is to be scheduled for a target TTI.
- the set of RBs or RBGs is shared by a set of MU -Ml MO layers supported by a network node in the wireless communication network.
- Each MU-MIMO layer of the set of MU-MIMO layers is assigned to a different UE of the list of UEs.
- the apparatus further comprises a means for obtaining a data vector for each RB or RBG of the set of RBs or RBGs for each MU-MIMO layer of the set of MU-MIMO layers.
- the data vector comprises frequency-specific information and a spatial correlation value.
- the frequency-specific information indicates a channel state of each UE of the list of UEs.
- the spatial correlation value indicates whether the UEs of the list of UEs are spatially separable by using beamforming.
- the apparatus further comprises a means for scheduling the set of RBs or RBGs for the list of UEs for the target TTI by using a ML model.
- the ML model comprises at least one NN configured to receive the data vectors as input data and output a scheduling decision sequentially for each RB or RBG of the set of RBs or RBGs.
- the scheduling decision indicates whether the RB or RBG is to be scheduled to one or more UEs of the list of UEs.
- the apparatus may take MU-MIMO aspects (i.e., the spatial diversity of the UEs) into account when scheduling (both DL and UL) RBs or RBGs for the UEs. Furthermore, the apparatus thus configured may perform the frequency and spatial domain RB or RBG scheduling regardless of: a maximum number of UEs to be scheduled to RBs or RBGs, a physical layer numerology/bandwidth (i.e., a number of RBs or RBGs per TTI), (digita l/ana log/hybrid) conventional and/or ML-based beamformers used in the MU-MIMO network, and a number of RF chains determining a maximum number of spatially scheduled UEs.
- a maximum number of UEs to be scheduled to RBs or RBGs i.e., a number of RBs or RBGs per TTI
- digita l/ana log/hybrid conventional and/or ML-based beamformers used in the MU
- FIG. 1 shows a block diagram of a wireless communication system in accordance with the prior art
- FIG. 2 shows a block diagram of a Resource Block (RB) scheduling apparatus in a wireless communication network in accordance with one example embodiment
- FIG. 3 shows a flowchart of a method for operating the apparatus of FIG. 2 in accordance with one example embodiment
- FIG. 4 explains how a Neural Network (NN) used in the method of FIG. 3 may process data vectors obtained for RBs or RBGs for each MU -Ml MO layer to make RB or RBG scheduling decisions;
- NN Neural Network
- FIG. 5 shows an exemplary RB scheduling table comprising the scheduling decisions made by the NN used in the method of FIG. 3 and rewards for the scheduling decisions;
- FIG. 6 shows two possible options how a processor included in the apparatus of FIG. 2 may construct tuples to be fed to the NN for its training;
- FIG. 7 shows a Cumulative Distribution Function (CDF) versus the CPU execution time required for doing the scheduling decisions per TTI in a system level simulator for two cases: the frequency selective even RB scheduling (dashed curve) and the RB scheduling based on the method of FIG. 3 (solid curve);
- CDF Cumulative Distribution Function
- FIGs. 8A and 8B show a throughput performance of Spatial Domain (SD) scheduling for a UE throughput and a cell throughput, respectively;
- SD Spatial Domain
- FIG. 9 shows a reward evolution with 1000 step observation windows during the first 300000 simulation steps equal to ⁇ 21 seconds in real time.
- FIGs. lOA and 10B show, respectively, a DL windowed UE throughput and first MU-MIMO layer scheduling CPU time which are obtained by using the apparatus of FIG. 2 in accordance with the method of FIG. 3 and the prior art Proportional Fair (PF) FD scheduler with a decoupled greedy SD scheduler.
- PF Proportional Fair
- a User Equipment may refer to an electronic computing device that is configured to perform wireless communications.
- the UE may be implemented as a mobile station, a mobile terminal, a mobile subscriber unit, a mobile phone, a cellular phone, a smart phone, a cordless phone, a personal digital assistant (PDA), a wireless communication device, a desktop computer, a laptop computer, a tablet computer, a gaming device, a netbook, a smartbook, an ultrabook, a medical mobile device or equipment, a biometric sensor, a wearable device (e.g., a smart watch, smart glasses, a smart wrist band, etc.), an entertainment device (e.g., an audio player, a video player, etc.), a vehicular component or sensor (e.g., a driver-assistance system), a smart meter/sensor, an unmanned vehicle (e.g., an industrial robot, a quadcopter, etc.) and its component (e.g., a selfd
- an unmanned vehicle e.g
- a network node may refer to a node in any of a Radio Access Network (RAN) and a Core Network (CN). It should be noted that the CN may refer to a network intended for connecting different RAN nodes by providing proper interfaces therebetween. The CN may also provide a gateway to other networks, for example, a Data Network (DN).
- RAN Radio Access Network
- CN Core Network
- DN Data Network
- the network node may be implemented as a fixed point of communication/communication node for a UE in a particular wireless communication network. More specifically, the RAN node may be used to connect the UE to the DN through the CN and may be referred to as a base transceiver station (BTS) in terms of the 2G communication technology, a NodeB in terms of the 3G communication technology, an evolved NodeB (eNodeB) in terms of the 4G communication technology, and a gNB in terms of the 5G New Radio (NR) communication technology.
- BTS base transceiver station
- NodeB in terms of the 3G communication technology
- eNodeB evolved NodeB
- gNB 5G New Radio
- the RAN node may serve different cells, such as a macrocell, a microcell, a picocell, a femtocell, and/or other types of cells.
- the macrocell may cover a relatively large geographic area (for example, at least several kilometers in radius).
- the microcell may cover a geographic area less than two kilometers in radius, for example.
- the picocell may cover a relatively small geographic area, such, for example, as offices, shopping malls, train stations, stock exchanges, etc.
- the femtocell may cover an even smaller geographic area (for example, a home).
- the network node may also refer to any of CN network functions, such as an Access and Mobility Management Function (AMF), a Session Management Function (SMF), Unified Data Management (UDM), User Plane Function (UPF), Policy Control Function (PCF), etc.
- AMF Access and Mobility Management Function
- SMF Session Management Function
- UDM Unified Data Management
- UPF User Plane Function
- PCF Policy Control Function
- the AMF supports termination of Non-Access Stratum (NAS) signalling, NAS ciphering and integrity protection, registration management, connection management, mobility management, access authentication and authorization, security context management.
- the SMF supports session management (session establishment, modification, release), UE IP address allocation and management, Dynamic Host Configuration Protocol (DHCP) functions, termination of NAS signalling related to the session management, downlink (DL) data notification, traffic steering configuration for the UPF for proper traffic routing.
- NAS Non-Access Stratum
- DHCP Dynamic Host Configuration Protocol
- UDM supports Authentication and Key Agreement (AKA) credentials generation, user identification handling, access authorization, subscription management.
- the UPF supports packet routing and forwarding, packet inspection, Quality of Service (QoS) handling, acts as an external Protocol Data Unit (PDU) session point of interconnect to the DN, and is an anchor point for intra- and inter- Radio Access Technology (RAT) mobility.
- the PCF supports a unified policy framework, providing policy rules to Control Plane (CP) functions, access subscription information for policy decisions in a Unified Data Repository (UDR).
- AKA Authentication and Key Agreement
- the UPF supports packet routing and forwarding, packet inspection, Quality of Service (QoS) handling, acts as an external Protocol Data Unit (PDU) session point of interconnect to the DN, and is an anchor point for intra- and inter- Radio Access Technology (RAT) mobility.
- the PCF supports a unified policy framework, providing policy rules to Control Plane (CP) functions, access subscription information for policy decisions in a Unified Data Repository (UDR).
- a wireless communication network in which one or more network nodes communicate with each other and/or with one or more UEs, may refer to a cellular or mobile network, a Wireless Local Area Network (WLAN), a Wireless Personal Area Networks (WPAN), a Wireless Wide Area Network (WWAN), a satellite communication (SATCOM) system, or any other type of wireless communication networks.
- WLAN Wireless Local Area Network
- WPAN Wireless Personal Area Networks
- WWAN Wireless Wide Area Network
- SATCOM satellite communication
- the cellular network may operate according to the Global System for Mobile Communications (GSM) standard, the Code-Division Multiple Access (CDMA) standard, the Wide-Band Code-Division Multiple Access (WCDM) standard, the Time-Division Multiple Access (TDMA) standard, or any other communication protocol standard
- GSM Global System for Mobile Communications
- CDMA Code-Division Multiple Access
- WDM Wide-Band Code-Division Multiple Access
- TDMA Time-Division Multiple Access
- the WLAN may operate according to one or more versions of the IEEE 802.11 standards
- the WPAN may operate according to the Infrared Data Association (IrDA), Wireless USB, Bluetooth, or ZigBee standard
- the WWAN may operate according to the Worldwide Interoperability for Microwave Access (WiMAX) standard.
- WiMAX Worldwide Interoperability for Microwave Access
- Multi-User Multiple Input Multiple Output may refer to a communication technology at which a plurality of UEs communicate with a network node by using the same time-frequency resources (i.e., resource blocks (RBs) or RB Groups (RBGs), and the network node considers communication signals from the UEs as MIMO signals and separates the signals accordingly.
- the MU-MIMO technology may be considered as Space-Division Multiple Access (SDMA) using a spatial channel as a resource in addition to the conventional time-frequency resources.
- SDMA Space-Division Multiple Access
- two or more UEs are appropriately (by specially designed DL and UL schedulers usually resided in a RAN node) selected to transmit and receive data at the same time by using the same RBs or RBGs.
- a large multi-user diversity effect can be obtained, and the cell capacity of the whole wireless communication system can be improved.
- FIG. 1 shows a block diagram of a wireless communication system 100 in accordance with the prior art.
- the system 100 comprises a (conventional or ML-based) receiver 102, a (conventional or ML-based) transmitter 104, and a ML-based RB scheduler 106.
- the ML-based RB scheduler 106 is typically provided in a RAN node (e.g., gNB).
- the receiver 102 and the transmitter 104 are assumed to be implemented in a single UE, and the ML-based RB scheduler 106 is assumed to communicate with them by using a Single-User MIMO (SU-MIMO) technology.
- SU-MIMO Single-User MIMO
- the ML-based RB scheduler 106 is configured to rule both the receiver 102 and the transmitter 104. In other words, the ML-based RB scheduler 106 can perform both DL and UL RB or RBG scheduling for any number of MIMO layers in the SU-MIMO network. However, if the SU-MIMO technology is replaced with a MU-MIMO technology in the system 100, the ML-based RB scheduler 106 will need to be further improved such that it is able to consider the MU-MIMO aspects (i.e., spatial diversity or spatial correlation) when performing the DL and UL RB or RBG scheduling for multiple UEs.
- the MU-MIMO aspects i.e., spatial diversity or spatial correlation
- the example embodiments disclosed herein provide a technical solution that allows RBs or RBGs to be scheduled in both frequency and spatial domains for UEs in a wireless communication network supporting the MU-MIMO technology.
- a list of UEs is obtained, for which a set of RBs or RBGs is to be scheduled for a target TTI.
- the set of RBs or RBGs is shared by a set of MU-MIMO layers each assigned to a different UE of the list of UEs.
- a data vector for each RB or RBG for each MU-MIMO layer is obtained.
- the data vector comprises frequency-specific information and a spatial correlation value.
- the frequency-specific information indicates a channel state of each UE of the list of UEs for the corresponding RB or RBG, while the spatial correlation value indicates whether the UEs of the lists of UEs are spatially separable for the corresponding RB or RBG by using beamforming.
- the set of RBs or RBGs is scheduled for the list of UEs for the target TTI by using an ML model.
- the ML model comprises at least one NN configured to receive the data vectors as input data and output a scheduling decision sequentially for each RB or RBG of the set of RBs or RBGs.
- FIG. 2 shows a block diagram of an RB scheduling apparatus 200 in a wireless communication network in accordance with one example embodiment.
- the apparatus 200 may be used instead of the ML-based RB scheduler 106 in the system 100 to take account of the MU -Ml MO aspects when performing the DL and UL RB or RBG scheduling.
- the apparatus 200 may be implemented either as part of a network node (RAN or CN node) or as an individual apparatus connected to the network node by means of wire or wirelessly.
- the apparatus 200 comprises a processor 202 and a memory 204.
- the memory 204 stores processor-executable instructions 206 which, when executed by the processor 202, cause the processor 202 to perform the aspects of the present disclosure, as will be described below in more detail.
- processor-executable instructions 206 which, when executed by the processor 202, cause the processor 202 to perform the aspects of the present disclosure, as will be described below in more detail.
- FIG. 2 the number, arrangement, and interconnection of the constructive elements constituting the apparatus 200, which are shown in FIG. 2, are not intended to be any limitation of the present disclosure, but merely used to provide a general idea of how the constructive elements may be implemented within the apparatus 200.
- the processor 202 may be replaced with several processors, as well as the memory 204 may be replaced with several removable and/or fixed storage devices, depending on particular applications.
- the processor 202 may perform different operations required to perform data reception and transmission, such, for example, as signal modulation/demodulation, encoding/decoding, etc.
- the apparatus 200 may further comprise an individual transceiver which can be configured to perform the required operations for data reception and transmission based on commands from the processor 202.
- the processor 202 may be implemented as a CPU, general-purpose processor, single-purpose processor, microcontroller, microprocessor, application specific integrated circuit (ASIC), field programmable gate array (FPGA), digital signal processor (DSP), complex programmable logic device, etc. It should be also noted that the processor 202 may be implemented as any combination of one or more of the aforesaid. As an example, the processor 202 may be a combination of two or more microprocessors.
- the memory 204 may be implemented as a classical nonvolatile or volatile memory used in the modern electronic computing machines.
- the nonvolatile memory may include Read-Only Memory (ROM), ferroelectric Random-Access Memory (RAM), Programmable ROM (PROM), Electrically Erasable PROM (EEPROM), solid state drive (SSD), flash memory, magnetic disk storage (such as hard drives and magnetic tapes), optical disc storage (such as CD, DVD and Blu-ray discs), etc.
- ROM Read-Only Memory
- RAM ferroelectric Random-Access Memory
- PROM Programmable ROM
- EEPROM Electrically Erasable PROM
- SSD solid state drive
- flash memory magnetic disk storage (such as hard drives and magnetic tapes), optical disc storage (such as CD, DVD and Blu-ray discs), etc.
- the volatile memory examples thereof include Dynamic RAM, Synchronous DRAM (SDRAM), Double Data Rate SDRAM (DDR SDRAM), Static RAM, etc.
- the processor-executable instructions 206 stored in the memory 204 may be configured as a computer-executable program code which causes the processor 202 to perform the aspects of the present disclosure.
- the computer-executable program code for carrying out operations or steps for the aspects of the present disclosure may be written in any combination of one or more programming languages, such as Java, C++, Python, or the like.
- the computer-executable program code may be in the form of a high-level language or in a precompiled form and be generated by an interpreter (also pre-stored in the memory 204) on the fly-
- FIG. 3 shows a flowchart of a method 300 for operating the apparatus 200 in accordance with one example embodiment.
- the method 300 starts with a step S302, in which the processor 202 obtains a list of UEs for which a set of RBs or RBGs is to be scheduled for a target TTI.
- the set of RBs or RBGs is shared by a set of MU-MIMO layers each assigned to a different UE of the list of UEs.
- the list of UEs may comprise all UEs that have or is expected to have data in a DL or UL transmission buffer, respectively.
- the list of UEs may be a shortlisted subset of all UEs with nonempty transmission buffers. Said subset may be generated by calculating scheduling priority metrics.
- Such metrics may be based on randomization, expected throughput, past average throughput, channel state information, or anything derived from these metrics.
- the list of UEs may be obtained based on a priority value assigned to each UE in the network. On top of that, the selection of UEs for which the set of RBs or RGs is to be scheduled may be selected randomly, by using a round robin principle, proportional-fair scheduling, Quality-of- Service (QoS) priority-based approach, or any combination thereof.
- the processor 202 may obtain the list of UEs by itself or, if the apparatus 200 is implemented as an individual apparatus, may receive it from a network node (e.g., a gNB).
- a network node e.g., a gNB
- the method 300 proceeds to a step S304, in which the processor 202 obtains a data vector for each RB or RBG of the set of RBs or RBGs for each MU-MIMO layer of the set of MU- MIMO layers.
- the data vector comprises frequency-specific information and a spatial correlation value.
- the frequency-specific information is assumed to indicate a channel state of each UE of the list of UEs, while the spatial correlation value is assumed to indicate whether the UEs of the list of UEs are spatially separable by using beamforming.
- the processor 202 may create the data vector by itself or may receive it from the network node.
- the spatial correlation value should hint towards the spatial correlation of UEs to be potentially spatially co-scheduled to the same RB or RBG with the already scheduled UEs on the same RB or RBG.
- the spatial correlation value may be represented by a Singular Value Decomposition (SVD) of combined precoders of the UEs of the list of UEs.
- SVD Singular Value Decomposition
- a default value of 0 or 1.0 may be used as the spatial correlation value.
- the SVD is calculated with already scheduled (for the same RB/RBG) UEs' precoders and a new UE whose input value is calculated.
- the spatial correlation value may be additionally or alternatively represented by a Power Normalization Loss (PLN) value calculated for each UE of the list of UEs.
- PPN Power Normalization Loss
- the spatial correlation value may be based on an estimated throughput calculated for each UE.
- the throughput estimation would increase computational complexity due to required MU-MIMO precoder calculations. If a certain UE of the list of UEs is already scheduled for a certain RB or RBG of the set of RBs or RBGs, all the inputs for that UE may be masked to zero for remaining the MU-MIMO layers for that particular RB or RBG.
- any other spatial correlation values may be used.
- the principle is that a single value should hint towards channel correlation between the already scheduled UEs and a new scheduling candidate UE.
- an estimation of a throughput increase caused by scheduling each UE to an additional MU-MIMO layer may be used as (part of) the spatial correlation value.
- the processor 202 has to calculate new MU- MIMO precoders (e.g., zero forcing precoders) for scheduling a candidate UE taking into account the already scheduled UEs for the RB or RBG and the candidate UE.
- throughput estimations can be derived for each UE for the RB/RBG.
- the same can also be calculated without new candidate UEs to have a value by which to compare whether the sum throughput increases or not by scheduling a new UE to an additional MU-MIMO layer. All of this may be done separately for each candidate UE to see, e.g., which scheduling selections increase the sum throughput estimate and which not.
- the frequency-specific information may comprise at least one of: Channel State Information (CSI) in the RB or RBG, a past average DL throughput in the RB or RBG, a past average UL throughput in the RB or RBG, a number of already scheduled other RBs or RBGs of the set of RBs or RBGs, an average size of transmitted DL packets in the RB or RBG, an average size of transmitted UL packets in the RB or RBG, a Modulation and Coding Scheme (MCS), and a scheduling decision made for the previous RB or RBG of the set of RBs or RBGs.
- CSI Channel State Information
- MCS Modulation and Coding Scheme
- the CSI may comprise at least one of a Radio Resource Management (RRM) measurement for the RB or RBG, a RRM measurement for a sub-band comprising the RB or RBG, and a RRM measurement averaged over an entire bandwidth comprising the sub-band.
- RRM Radio Resource Management
- the RRM measurement may comprise a SINR, a RSRP, a RSRQ, a CQI, and/or any other signal quality parameter.
- the knowledge of the scheduling decision made for the previous RB or RBG of the set of RBs or RBGs may help the processor 202 provide continuous RB or RBG allocations required for continuous waveforms.
- the method 300 goes on to a step S306, in which the processor 202 schedules the set of RBs or RBGs for the list of UEs for the target TTI by using an ML model.
- the ML model comprises at least one NN configured to receive the data vectors as input data and output a scheduling decision sequentially for each RB or RBG of the set of RBs or RBGs.
- the scheduling decision indicates whether the RB or RBG is to be scheduled for one or more UEs of the list of UEs.
- the NN(s) output(s) optimal UE selection for each RB or RBG and each MU-MIMO layer.
- FIG. 4 explains how the NN used in the method 300 may process the data vectors obtained for the RBs or RBGs for each MU-MIMO layer to make the scheduling decisions for the RBs or RBGs.
- FIG. 4 it is assumed that only one NN is used, which may process the data vectors one by one, thereby outputting the scheduling decisions for the RBs or RBGs one by one in the order of the RBs or the RBGs in each MU-MIMO layer (see the dashed line in FIG. 4). If a set of NNs is used with each of them being assigned to a different MU-MIMO layer, the scheduling decisions may be outputted for the RBs or RBGs in parallel.
- the NN used in the step S306 may be trained (in advance or online during the method 300) based on a sequence of tuples by using a reinforcement learning algorithm (e.g., a DDQN based reinforcement learning algorithm).
- a reinforcement learning algorithm e.g., a DDQN based reinforcement learning algorithm
- Each tuple of the sequence of tuples comprises: (a) a data vector obtained for a target RB or RBG of the set of RBs or RBGs for a target MU -Ml MO layer of the set of MU-MIMO layers for the target TTI; (b) a scheduling decision outputted by the NN for the target RB or RBG; (c) a performance reward resulted from the scheduling decision for the target RB or RBG; and (d) a data vector obtained for one of: (i) a subsequent RB or RBG of the set of RBs or RBGs for the target MU-MIMO layer for the target TTI, (ii) the target RB or RBG for a subsequent MU-MIMO layer of the set of MU- MIMO layers for the target TTI, and (iii) the target RB or RBG for the target MU-MIMO layer for a subsequent TTI.
- the ML model may output a table of all scheduling decisions each indicating N optimal UEs of the list of UEs for each RB or RBG for each MU-MIMO layer, where N is the maximum number of beams that may be formed/selected spatially for the same RB(s) or RBG(s).
- FIG. 5 shows an exemplary RB scheduling table 500 comprising the scheduling decisions made by the NN used in the step S306 of the method 300 and rewards for the scheduling decisions.
- each column corresponds to one MU-MIMO layer of the set of MU-MIMO layers, while each row corresponds to one RB of the set of RBs.
- the scheduling decision for UE1 consists in that each of the RBs should be scheduled for UE1 in the first MU- MIMO layer, the scheduling decision for UE2 consists in that only the first two RBs should be scheduled for UE2 in the second MU-MIMO layer, and the scheduling decision for UE3 consists in that only the second and last RBs should be scheduled for UE3 in the last MU-MIMO layer.
- the table 500 may also comprise common performance rewards maximized for each RB. In other words, the same performance reward may be used for all the MU-MIMO layers for a single RB, regardless of whether the NN has decided to allocate it or not.
- the reward function should sum the bytes per RB and use the max past average throughput of the coscheduled UEs on the same RB, namely: where is the past average throughput (bytes per RB) of /-th UE. This way, the proportional fairness is ensured, and the MU-MIMO gain always increases the reward.
- the processor 202 may collect the above-mentioned tuples and train the NN based by using the tuples one by one.
- the reward may be based on an expert scheduler, which may be any well performing heuristic scheduler algorithm.
- the NN may be trained to mimic such existing expert schedulers by rewarding a positive reward if the scheduling decision given by the NN is the same as that provided by the expert scheduler. Accordingly, a negative reward may be given if the scheduling decision does not match the expert decision.
- the scheduling logic of the expert scheduler may be transferred into the ML model. The benefit of this is to achieve a much less computationally complex scheduler that may at least match the performance of the expert scheduler.
- FIG. 6 shows two possible options how the processor 202 may construct the tuples to be fed to the NN for its training.
- the ML model may use the so-called replay memory for storing the tuples. If the first and last elements in each tuple (i.e., elements (a) and (d)) correspond, respectively, to the previous and next (in sequence) RBs for the same MU-MIMO layer, then the NN may be trained by performing seamless state transitions over all RBs in one MU-MIMO (as schematically shown by using the vertical arrows in FIG. 6).
- the NN may be trained by performing seamless state transitions for the same RB over all the MU-MIMO layers (as schematically shown by using the horizontal arrows in FIG. 6). It should be also noted that the present disclosure is not limited to these two options shown in FIG.
- the first and last elements in each tuple may correspond to the same RB in the same MU-MIMO layer but in the previous and next (in sequence) TTIs (this will imply the combination of this option with one of those shown in FIG. 6).
- a fully digital beamformer using regularized zero forcing (ZF) as a beamformer was assumed.
- ZF regularized zero forcing
- a dynamic system level simulator with state-of-the-art MU-MIMO baseline parameterization was used for benchmarking the ML-based scheduling solution.
- the maximum number of MU-MIMO layers in the system level simulator was set to 4 in a macro cell simulation scenario.
- FIG. 7 shows a Cumulative Distribution Function (CDF) versus the CPU execution time required for doing the scheduling decisions per TTI for two cases: the frequency selective even RB scheduling (dashed curve) and the RB scheduling based on the method 300 (solid curve).
- CDF Cumulative Distribution Function
- FIG. 7 shows absolute dominance in the real-time scheduling execution time. The execution times are shown for a single downlink MU-MIMO layer scheduling all RBs. The difference would be even more ludicrous if the prior art heuristic SD scheduler would be taken into account in total execution time or the more complex prior art heuristic FD scheduler would be used.
- "Even RB scheduling” is one of the simplest frequency selective scheduling algorithms, and still the method 300 can outperform it in the execution time due to the lack of sub-band channel state information processing, downlink throughput estimation calculations for each UE for each RB.
- FIGs. 8A and 8B show a throughput performance of SD scheduling for a UE throughput and a cell throughput, respectively.
- the UE was set as an invalid candidate and UE's input values to be fed to the NN were masked with zeros once a number of scheduling candidates per total number of available RBs were reached for the UE. If the NN used in the method 300 picks an invalid candidate UE, it generates an empty allocation for the RB for the current MU-MIMO layer.
- FIG. 9 shows a reward evolution with 1000 step observation windows during the first 300000 simulation steps equal to ⁇ 21 seconds in real time.
- the processor 202 has collected and trained the NN with bit more than 10000 tuples (as described above).
- the simulation was started with 100% exploration rate. In other words, all the scheduling decisions are randomly picked.
- the exploration rate is linearly decreased to zero before 75000 steps.
- the NN is considered to be enough stabilized for the start of the actual simulation and its result collection. Whether the training was stopped or continued after that point did not much affect to the simulation results.
- FIGs. lOA and 10B show, respectively, a DL windowed UE throughput and first MU- Ml MO layer scheduling CPU time which are obtained by using the apparatus 200 in accordance with the method 300 and the prior art Proportional Fair (PF) FD scheduler with a decoupled greedy SD scheduler.
- PF Proportional Fair
- FIGs. lOA and 10B show, respectively, a DL windowed UE throughput and first MU- Ml MO layer scheduling CPU time which are obtained by using the apparatus 200 in accordance with the method 300 and the prior art Proportional Fair (PF) FD scheduler with a decoupled greedy SD scheduler.
- PF Proportional Fair
- the scheduling complexity reduction of all MU-MIMO layers is in the same ballpark, because the apparatus 200 does the same for all MU-MIMO layers, whereas the heuristic SD scheduler keeps re-calculating precoders, re-calculating throughput estimations, reselecting MCSs, etc. (i.e., everything that the apparatus 200 does only once after the scheduling decisions).
- each step or operation of the method 300 can be implemented by various means, such as hardware, firmware, and/or software.
- one or more of the steps or operations described above can be embodied by processor executable instructions, data structures, program modules, and other suitable data representations.
- the processor-executable instructions which embody the steps or operations described above can be stored on a corresponding data carrier and executed by the processor 202.
- This data carrier can be implemented as any computer-readable storage medium configured to be readable by said at least one processor to execute the processor executable instructions.
- Such computer-readable storage media can include both volatile and nonvolatile media, removable and non-removable media.
- the computer-readable media comprise media implemented in any method or technology suitable for storing information.
- the practical examples of the computer-readable media include, but are not limited to information-delivery media, RAM, ROM, EEPROM, flash memory or other memory technology, CD-ROM, digital versatile discs (DVD), holographic media orotheroptical disc storage, magnetictape, magnetic cassettes, magnetic disk storage, and other magnetic storage devices.
Landscapes
- Engineering & Computer Science (AREA)
- Computer Networks & Wireless Communication (AREA)
- Signal Processing (AREA)
- Radio Transmission System (AREA)
- Mobile Radio Communication Systems (AREA)
Abstract
Description
Claims
Priority Applications (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| PCT/EP2023/066263 WO2024256023A1 (en) | 2023-06-16 | 2023-06-16 | Resource block scheduling in wireless communication network |
| EP23734495.7A EP4728656A1 (en) | 2023-06-16 | 2023-06-16 | Resource block scheduling in wireless communication network |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| PCT/EP2023/066263 WO2024256023A1 (en) | 2023-06-16 | 2023-06-16 | Resource block scheduling in wireless communication network |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2024256023A1 true WO2024256023A1 (en) | 2024-12-19 |
Family
ID=87059747
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/EP2023/066263 Ceased WO2024256023A1 (en) | 2023-06-16 | 2023-06-16 | Resource block scheduling in wireless communication network |
Country Status (2)
| Country | Link |
|---|---|
| EP (1) | EP4728656A1 (en) |
| WO (1) | WO2024256023A1 (en) |
Citations (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20230189317A1 (en) * | 2021-12-15 | 2023-06-15 | Intel Corporation | User scheduling using a graph neural network |
-
2023
- 2023-06-16 EP EP23734495.7A patent/EP4728656A1/en active Pending
- 2023-06-16 WO PCT/EP2023/066263 patent/WO2024256023A1/en not_active Ceased
Patent Citations (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20230189317A1 (en) * | 2021-12-15 | 2023-06-15 | Intel Corporation | User scheduling using a graph neural network |
Also Published As
| Publication number | Publication date |
|---|---|
| EP4728656A1 (en) | 2026-04-22 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US11212858B2 (en) | On-demand backhaul link management measurements for integrated access backhaul for 5G or other next generation network | |
| US10833751B2 (en) | Facilitation of user equipment specific compression of beamforming coefficients for fronthaul links for 5G or other next generation network | |
| US10827547B2 (en) | Radio resource configuration and measurements for integrated access backhaul for 5G or other next generation network | |
| US10826578B2 (en) | Facilitation of beamforming gains for fronthaul links for 5G or other next generation network | |
| JP2020502878A (en) | Asynchronous multipoint transmission method | |
| US11949505B2 (en) | Scheduling of uplink data using demodulation reference signal and scheduled resources | |
| EP3085183A1 (en) | A network node and method for enabling interference alignment of transmissions to user equipments | |
| US20200351746A1 (en) | Mitigating user equipment overheating for 5g or other next generation network | |
| US11165475B2 (en) | Linear combination codebook based per layer power allocation feedback for 5G or other next generation network | |
| US12225412B2 (en) | Adaptive radio access network bit rate scheduling | |
| US11038626B2 (en) | Hybrid automatic repeat request reliability for 5G or other next generation network | |
| US11075676B2 (en) | Facilitating semi-open loop based transmission diversity for uplink transmissions for 5G or other next generation networks | |
| US20240397375A1 (en) | Flexible configuration of guaranteed bitrate admission control for 5g or other next generation network | |
| CN102300326B (en) | Scheduling method of multi-user multi-input multi-output (MIMO) communication system and base station | |
| EP4728656A1 (en) | Resource block scheduling in wireless communication network | |
| US9948376B1 (en) | Transmission mode selection | |
| WO2025171858A1 (en) | Machine learning-based resource block scheduling in multi-user multiple-input multiple-output network |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 23734495 Country of ref document: EP Kind code of ref document: A1 |
|
| WWE | Wipo information: entry into national phase |
Ref document number: 202647001879 Country of ref document: IN |
|
| WWE | Wipo information: entry into national phase |
Ref document number: 2023734495 Country of ref document: EP |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| ENP | Entry into the national phase |
Ref document number: 2023734495 Country of ref document: EP Effective date: 20260116 |
|
| WWP | Wipo information: published in national office |
Ref document number: 202647001879 Country of ref document: IN |
|
| ENP | Entry into the national phase |
Ref document number: 2023734495 Country of ref document: EP Effective date: 20260116 |
|
| ENP | Entry into the national phase |
Ref document number: 2023734495 Country of ref document: EP Effective date: 20260116 |
|
| ENP | Entry into the national phase |
Ref document number: 2023734495 Country of ref document: EP Effective date: 20260116 |
|
| WWP | Wipo information: published in national office |
Ref document number: 2023734495 Country of ref document: EP |

