Reinforcement Learning: An Introduction
Data up to Jan 2025
Total Citations Per Year
Abstract
References (80)
Genetic algorithms in search, optimization, and machine learning
1989 • 49,726 citations
Dynamic Programming
1957 • 13,549 citations
Pattern classification and scene analysis
1973 • 12,976 citations
Matrix Iterative Analysis
2000 • 5,826 citations
Learning from delayed rewards
1989 • 5,606 citations
Parallel and Distributed Computation: Numerical Methods
1989 • 5,603 citations
Theory and Practice of Recursive Identification
1983 • 4,095 citations
Untitled
1988 • 3,901 citations
Some Studies in Machine Learning Using the Game of Checkers
1959 • 3,897 citations
Dynamic Programming and Markov Processes.
1961 • 3,612 citations
Adaptive filtering prediction and control
1984 • 3,449 citations
Radial Basis Functions, Multi-Variable Functional Interpolation and Adaptive Networks
1988 • 3,440 citations
The Behavior of Organisms
1938 • 3,256 citations
Neuronlike adaptive elements that can solve difficult learning control problems
1983 • 3,122 citations
Purposive Behavior in Animals and Men
1950 • 2,502 citations
Some aspects of the sequential design of experiments
1952 • 2,215 citations
ON THE LIKELIHOOD THAT ONE UNKNOWN PROBABILITY EXCEEDS ANOTHER IN VIEW OF THE EVIDENCE OF TWO SAMPLES
1933 • 2,206 citations
The nature of explanation
1943 • 1,734 citations
Learning Automata: An Introduction
1989 • 1,597 citations
Stochastic models for learning.
1955 • 1,540 citations
Toward a modern theory of adaptive networks: Expectation and prediction.
1981 • 1,416 citations
Introduction to Stochastic Dynamic Programming.
1986 • 1,200 citations
Radial basis functions for multivariable interpolation: a review
1987 • 1,170 citations
Stochastic systems : estimation, identification, and adaptive control
1986 • 1,049 citations
Sparse Distributed Memory
1988 • 993 citations
Toward a statistical theory of learning.
1950 • 980 citations
Distinctive features, categorical perception, and probability learning: Some applications of a neural model.
1977 • 875 citations
Temporal credit assignment in reinforcement learning
1984 • 746 citations
A Theory of Networks for Approximation and Learning
1989 • 721 citations
Adaptive Behavior and Learning
1985 • 716 citations
Monte Carlo methods
1986 • 694 citations
Learning Automata - A Survey
1974 • 668 citations
The Art and Theory of Dynamic Programming
1977 • 498 citations
What are plans for?
1990 • 378 citations
Is there a cell-biological alphabet for simple forms of learning?
1984 • 334 citations
Pattern-recognizing stochastic learning automata
1985 • 309 citations
Punish/Reward: Learning with a Critic in Adaptive Threshold Systems
1973 • 279 citations
Adaptive Treatment Allocation and the Multi-Armed Bandit Problem
1987 • 277 citations
Simulation of self-organizing systems by digital computer
1954 • 259 citations
Modified Policy Iteration Algorithms for Discounted Markov Decision Problems
1978 • 239 citations
A Survey of Some Results in Stochastic Adaptive Control
1985 • 227 citations
Learning control systems--Review and outlook
1970 • 217 citations
An adaptive optimal controller for discrete-time Markov environments
1977 • 179 citations
A new approach to the design of reinforcement schemes for learning automata
1985 • 175 citations
Intelligent behavior as an adaptation to the task environment
1982 • 172 citations
Landmark learning: An illustration of associative search
1981 • 151 citations
Connectionist learning for control: an overview
1990 • 130 citations
On the Theory of Apportionment
1935 • 129 citations
Functional approximations and dynamic programming
1959 • 126 citations
Foundations of conditioning and learning
1967 • 119 citations
Polynomial approximation—a new computational technique in dynamic programming: Allocation processes
1963 • 115 citations
Learning and Problem Solving with Multilayer Connectionist Systems
1986 • 114 citations
8 Reinforcement-Learning Control and Pattern Recognition Systems
1970 • 109 citations
Simulation of anticipatory responses in classical conditioning by a neuron-like adaptive element
1982 • 104 citations
Simulation of the classically conditioned nictitating membrane response by a neuron-like adaptive element: Response topography, neuronal firing, and interstimulus intervals
1986 • 87 citations
A Chess-Playing Machine
1950 • 83 citations
Programming a Computer for Playing Chess
1988 • 81 citations
Further Real Applications of Markov Decision Processes
1988 • 80 citations
On the use of backpropagation in associative reinforcement learning
1988 • 72 citations
Brownian Motion and Potential Theory
1969 • 68 citations
The Logic of Limax Learning
1985 • 57 citations
Diversity-based inference of finite automata
1987 • 56 citations
Beat the Dealer: A Winning Strategy for the Game of Twenty-One
1964 • 52 citations
ASSOCIATIVE SEARCH NETWORK - A REINFORCEMENT LEARNING ASSOCIATIVE MEMORY
1981 • 50 citations
A unified theory of heuristic evaluation functions and its application to learning
1986 • 40 citations
Generalization of pattern recognition in a self-organizing system
1955 • 39 citations
On thought: the extrinsic theory.
1956 • 37 citations
Splines and efficiency in dynamic programming
1976 • 35 citations
A comparison and evaluation of three machine learning procedures as applied to the game of checkers
1974 • 29 citations
The CDP: A Unifying Formulation for Heuristic Search, Dynamic Programming, and Branch-and-Bound
1988 • 29 citations
Signature Table Systems and Learning
1982 • 25 citations
A learning machine with monologue
1969 • 24 citations
STELLA: A scheme for a learning machine
1963 • 24 citations
A neural model of adaptive behavior
1983 • 16 citations
Human operators and automatic adaptive controllers: A comparative study on a particular control task
1973 • 13 citations
Applications of artificial intelligence techniques to a spacecraft control problem
1967 • 12 citations
Computational Capabilities of Single Neurons: Relationship to Simple Forms of Associative and Nonassociative Learning in Aplysia
1989 • 10 citations
Game-theoretic cooperativity in networks of self-interested units
1987 • 5 citations
A New Machine-Learning Technique Applied to the Game of Checkers
1966 • 5 citations
Deleted Work
1955 • 0 citations
Cited By (0)
No citing papers found in database