Reinforcement learning ramp metering without complete information (Q763169): Difference between revisions
From MaRDI portal
Created a new Item |
ReferenceBot (talk | contribs) Changed an Item |
||
(4 intermediate revisions by 4 users not shown) | |||
Property / Wikidata QID | |||
Property / Wikidata QID: Q58907923 / rank | |||
Normal rank | |||
Property / MaRDI profile type | |||
Property / MaRDI profile type: MaRDI publication profile / rank | |||
Normal rank | |||
Property / full work available at URL | |||
Property / full work available at URL: https://doi.org/10.1155/2012/208456 / rank | |||
Normal rank | |||
Property / OpenAlex ID | |||
Property / OpenAlex ID: W1976365331 / rank | |||
Normal rank | |||
Property / cites work | |||
Property / cites work: Q4517722 / rank | |||
Normal rank | |||
Property / cites work | |||
Property / cites work: Micro- and macro-simulation of freeway traffic / rank | |||
Normal rank | |||
Property / cites work | |||
Property / cites work: Synchronized flow as a new traffic phase and related problems for traffic flow modelling / rank | |||
Normal rank | |||
links / mardi / name | links / mardi / name | ||
Latest revision as of 00:14, 5 July 2024
scientific article
Language | Label | Description | Also known as |
---|---|---|---|
English | Reinforcement learning ramp metering without complete information |
scientific article |
Statements
Reinforcement learning ramp metering without complete information (English)
0 references
9 March 2012
0 references
Summary: This paper develops a model of reinforcement learning ramp metering (RLRM) without complete information, which is applied to alleviate traffic congestions on ramps. RLRM consists of prediction tools depending on traffic flow simulation and optimal choice model based on reinforcement learning theories. Moreover, it is also a dynamic process with abilities of automaticity, memory and performance feedback. Numerical cases are given in this study to demonstrate RLRM such as calculating outflow rate, density, average speed, and travel time compared to no control and fixed-time control. Results indicate that the greater is the inflow, the more is the effect. In addition, the stability of RLRM is better than fixed-time control.
0 references
reinforcement learning ramp metering (RLRM)
0 references
traffic flow simulation
0 references
optimal choice model
0 references
reinforcement learning theories
0 references
stability of RLRM
0 references