EP4711983A2 - Verfahren und system für multibatch-verstärkungslernen über multiimitationslernen - Google Patents
Verfahren und system für multibatch-verstärkungslernen über multiimitationslernenInfo
- Publication number
- EP4711983A2 EP4711983A2 EP25224666.5A EP25224666A EP4711983A2 EP 4711983 A2 EP4711983 A2 EP 4711983A2 EP 25224666 A EP25224666 A EP 25224666A EP 4711983 A2 EP4711983 A2 EP 4711983A2
- Authority
- EP
- European Patent Office
- Prior art keywords
- base station
- state
- cell
- traffic data
- traffic
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Pending
Links
Classifications
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04W—WIRELESS COMMUNICATION NETWORKS
- H04W28/00—Network traffic management; Network resource management
- H04W28/02—Traffic management, e.g. flow control or congestion control
- H04W28/08—Load balancing or load distribution
- H04W28/086—Load balancing or load distribution among access entities
- H04W28/0861—Load balancing or load distribution among access entities between base stations
- H04W28/0862—Load balancing or load distribution among access entities between base stations of same hierarchy level
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04L—TRANSMISSION OF DIGITAL INFORMATION, e.g. TELEGRAPHIC COMMUNICATION
- H04L41/00—Arrangements for maintenance, administration or management of data switching networks, e.g. of packet switching networks
- H04L41/16—Arrangements for maintenance, administration or management of data switching networks, e.g. of packet switching networks using machine learning or artificial intelligence
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04W—WIRELESS COMMUNICATION NETWORKS
- H04W24/00—Supervisory, monitoring or testing arrangements
- H04W24/02—Arrangements for optimising operational condition
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06N—COMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N20/00—Machine learning
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06N—COMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computing arrangements based on biological models
- G06N3/02—Neural networks
- G06N3/08—Learning methods
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04W—WIRELESS COMMUNICATION NETWORKS
- H04W24/00—Supervisory, monitoring or testing arrangements
- H04W24/10—Scheduling measurement reports ; Arrangements for measurement reports
Landscapes
- Engineering & Computer Science (AREA)
- Computer Networks & Wireless Communication (AREA)
- Signal Processing (AREA)
- Artificial Intelligence (AREA)
- Computer Vision & Pattern Recognition (AREA)
- Databases & Information Systems (AREA)
- Evolutionary Computation (AREA)
- Medical Informatics (AREA)
- Software Systems (AREA)
- Mobile Radio Communication Systems (AREA)
Applications Claiming Priority (5)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US202163253023P | 2021-10-06 | 2021-10-06 | |
| US202163253823P | 2021-10-08 | 2021-10-08 | |
| US17/957,960 US12452733B2 (en) | 2021-10-06 | 2022-09-30 | Multi-batch reinforcement learning via multi-imitation learning |
| PCT/KR2022/015057 WO2023059106A1 (en) | 2021-10-06 | 2022-10-06 | Method and system for multi-batch reinforcement learning via multi-imitation learning |
| EP22878927.7A EP4324241B1 (de) | 2021-10-06 | 2022-10-06 | Verfahren und system für multibatch-verstärkungslernen über multiimitationslernen |
Related Parent Applications (2)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP22878927.7A Division EP4324241B1 (de) | 2021-10-06 | 2022-10-06 | Verfahren und system für multibatch-verstärkungslernen über multiimitationslernen |
| EP22878927.7A Division-Into EP4324241B1 (de) | 2021-10-06 | 2022-10-06 | Verfahren und system für multibatch-verstärkungslernen über multiimitationslernen |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| EP4711983A2 true EP4711983A2 (de) | 2026-03-18 |
| EP4711983A3 EP4711983A3 (de) | 2026-04-08 |
Family
ID=85774948
Family Applications (2)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP25224666.5A Pending EP4711983A3 (de) | 2021-10-06 | 2022-10-06 | Verfahren und system für multibatch-verstärkungslernen über multiimitationslernen |
| EP22878927.7A Active EP4324241B1 (de) | 2021-10-06 | 2022-10-06 | Verfahren und system für multibatch-verstärkungslernen über multiimitationslernen |
Family Applications After (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP22878927.7A Active EP4324241B1 (de) | 2021-10-06 | 2022-10-06 | Verfahren und system für multibatch-verstärkungslernen über multiimitationslernen |
Country Status (4)
| Country | Link |
|---|---|
| US (2) | US12452733B2 (de) |
| EP (2) | EP4711983A3 (de) |
| CN (1) | CN117941413A (de) |
| WO (1) | WO2023059106A1 (de) |
Families Citing this family (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US11856524B2 (en) | 2021-12-07 | 2023-12-26 | Intelligent Automation, Llc | Systems and methods for deep reinforcement learning for optimal power control in wireless networks |
| CN118780387A (zh) * | 2023-04-10 | 2024-10-15 | 华为技术有限公司 | 用于训练决策模型的方法、装置、设备、介质和程序产品 |
Family Cites Families (12)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN105072638B (zh) | 2015-07-28 | 2018-11-09 | 北京邮电大学 | 一种时延忍受异构无线移动网络的流量分载方法 |
| US10789544B2 (en) | 2016-04-05 | 2020-09-29 | Google Llc | Batching inputs to a machine learning model |
| CN108848520B (zh) | 2018-05-28 | 2020-04-10 | 西安交通大学 | 一种基于流量预测与基站状态的基站休眠方法 |
| US11134016B2 (en) | 2018-10-26 | 2021-09-28 | Hughes Network Systems, Llc | Monitoring a communication network |
| US11044185B2 (en) | 2018-12-14 | 2021-06-22 | At&T Intellectual Property I, L.P. | Latency prediction and guidance in wireless communication systems |
| US12262240B2 (en) | 2020-01-08 | 2025-03-25 | Telefonaktiebolaget Lm Ericsson (Publ) | Methods for intelligent resource allocation based on throttling of user equipment traffic and related apparatus |
| US20210289586A1 (en) | 2020-03-10 | 2021-09-16 | Assia Spe, Llc | System and method for closed loop automation between wifi wireless network nodes |
| CN113498076B (zh) | 2020-03-20 | 2026-01-13 | 北京三星通信技术研究有限公司 | 基于o-ran的性能优化配置方法与设备 |
| US12561602B2 (en) | 2020-06-17 | 2026-02-24 | Toyota Research Institute, Inc. | Reinforcement learning based control of imitative policies for autonomous driving |
| CN112884075A (zh) | 2021-03-23 | 2021-06-01 | 北京天融信网络安全技术有限公司 | 一种流量数据增强方法、流量数据分类方法及相关装置 |
| CN113365312B (zh) | 2021-06-22 | 2022-10-14 | 东南大学 | 强化学习和监督学习相结合的移动负载均衡方法 |
| EP4175370A3 (de) * | 2021-10-28 | 2023-08-30 | Nokia Solutions and Networks Oy | Energieeinsparung in einem funkzugangsnetzwerk |
-
2022
- 2022-09-30 US US17/957,960 patent/US12452733B2/en active Active
- 2022-10-06 WO PCT/KR2022/015057 patent/WO2023059106A1/en not_active Ceased
- 2022-10-06 EP EP25224666.5A patent/EP4711983A3/de active Pending
- 2022-10-06 CN CN202280057977.XA patent/CN117941413A/zh active Pending
- 2022-10-06 EP EP22878927.7A patent/EP4324241B1/de active Active
-
2025
- 2025-09-22 US US19/335,545 patent/US20260012850A1/en active Pending
Also Published As
| Publication number | Publication date |
|---|---|
| CN117941413A (zh) | 2024-04-26 |
| EP4324241C0 (de) | 2026-01-28 |
| EP4711983A3 (de) | 2026-04-08 |
| WO2023059106A1 (en) | 2023-04-13 |
| US12452733B2 (en) | 2025-10-21 |
| EP4324241A1 (de) | 2024-02-21 |
| EP4324241A4 (de) | 2024-10-23 |
| US20230107539A1 (en) | 2023-04-06 |
| EP4324241B1 (de) | 2026-01-28 |
| US20260012850A1 (en) | 2026-01-08 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| Chen et al. | Dynamic task allocation and service migration in edge-cloud IoT system based on deep reinforcement learning | |
| Wei et al. | Joint optimization of caching, computing, and radio resources for fog-enabled IoT using natural actor–critic deep reinforcement learning | |
| Sonmez et al. | Machine learning-based workload orchestrator for vehicular edge computing | |
| US11388732B2 (en) | Method for associating user equipment in a cellular network via multi-agent reinforcement learning | |
| Luong et al. | Deep reinforcement learning-based resource allocation in cooperative UAV-assisted wireless networks | |
| US20250324325A1 (en) | Method of load forecasting via knowledge distillation, and an apparatus for the same | |
| CN111967605B (zh) | 无线电接入网中的机器学习 | |
| CN116406004B (zh) | 无线网络资源分配系统的构建方法和资源管理方法 | |
| US20260012850A1 (en) | Multi-batch reinforcement learning via multi-imitation learning | |
| US20250168255A1 (en) | Method of performing communication load balancing with multi-teacher reinforcement learning, and an apparatus for the same | |
| Cao et al. | Deep reinforcement learning for multi-user access control in UAV networks | |
| Feng et al. | Goal-oriented wireless communication resource allocation for cyber-physical systems | |
| Galli et al. | Playing with a multi armed bandit to optimize resource allocation in satellite-enabled 5g networks | |
| Khoramnejad et al. | Carrier aggregation, load balancing, and backhauling in non-terrestrial networks: Generative diffusion model-based optimization | |
| Sun et al. | Hierarchical reinforcement learning for AP duplex mode optimization in network-assisted full-duplex cell-free networks | |
| Liu et al. | Toward mobility-aware edge inference via model partition and service migration | |
| Hu et al. | Inter-cell network slicing with transfer learning empowered multi-agent deep reinforcement learning | |
| Lin et al. | Online task offloading in udn: A deep reinforcement learning approach with incomplete information | |
| Majumdar et al. | Improving scalability of 6G network automation with distributed deep Q-networks | |
| Zhu et al. | Online function scheduling for dual-heterogeneous serverless vehicular edge computing | |
| Zhang et al. | Cooperative optimisation strategy of computation offloading in multi‐UAVs‐assisted edge computing networks | |
| Wu et al. | Joint optimization of flying trajectory and task offloading for UAV-enabled MEC networks: A digital twin-assisted hybrid learning approach | |
| Chu et al. | Reinforcement learning based multi-access control with energy harvesting | |
| WO2024147107A1 (en) | Using inverse reinforcement learning in objective-aware traffic flow prediction | |
| Wang et al. | Optimizing Proximity Strategy for Federated Learning Node Selection in the Space-Air–Ground Information Network for Smart Cities |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| PUAI | Public reference made under article 153(3) epc to a published international application that has entered the european phase |
Free format text: ORIGINAL CODE: 0009012 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE APPLICATION HAS BEEN PUBLISHED |
|
| REG | Reference to a national code |
Ref country code: DE Ref legal event code: R079 Free format text: PREVIOUS MAIN CLASS: G06N0003080000 Ipc: H04W0028080000 |
|
| PUAL | Search report despatched |
Free format text: ORIGINAL CODE: 0009013 |
|
| AC | Divisional application: reference to earlier application |
Ref document number: 4324241 Country of ref document: EP Kind code of ref document: P |
|
| AK | Designated contracting states |
Kind code of ref document: A2 Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC ME MK MT NL NO PL PT RO RS SE SI SK SM TR |
|
| AK | Designated contracting states |
Kind code of ref document: A3 Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC ME MK MT NL NO PL PT RO RS SE SI SK SM TR |
|
| RIC1 | Information provided on ipc code assigned before grant |
Ipc: H04W 28/08 20230101AFI20260303BHEP Ipc: G06N 3/04 20230101ALI20260303BHEP Ipc: G06N 3/08 20230101ALI20260303BHEP |