Mittal et al., 2025 - Google Patents
Meta-Learning for Few-Shot Adaptation in Robotic Control TasksMittal et al., 2025
- Document ID
- 5697754103309029122
- Author
- Mittal S
- Makhmudov S
- Matniyozov S
- Matchanova B
- Sobirov A
- Eshchanov T
- Publication year
- Publication venue
- 2025 2nd International Conference on Recent Trends in Electrical, Electronics and Computing Technologies (ICRTEECT)
External Links
Snippet
Meta-learning is emerging as a powerful paradigm to allow robots to learn new tasks quickly and with little data, to address an obstacle in robot control. This paper investigates few-shot adaptation methods such as MAML, Reptile, and prototypical network that leverage …
Classifications
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING; COUNTING
- G06N—COMPUTER SYSTEMS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computer systems based on biological models
- G06N3/02—Computer systems based on biological models using neural network models
- G06N3/04—Architectures, e.g. interconnection topology
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING; COUNTING
- G06N—COMPUTER SYSTEMS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N99/00—Subject matter not provided for in other groups of this subclass
- G06N99/005—Learning machines, i.e. computer in which a programme is changed according to experience gained by the machine itself during a complete run
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING; COUNTING
- G06N—COMPUTER SYSTEMS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computer systems based on biological models
- G06N3/02—Computer systems based on biological models using neural network models
- G06N3/08—Learning methods
- G06N3/086—Learning methods using evolutionary programming, e.g. genetic algorithms
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING; COUNTING
- G06N—COMPUTER SYSTEMS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computer systems based on biological models
- G06N3/02—Computer systems based on biological models using neural network models
- G06N3/06—Physical realisation, i.e. hardware implementation of neural networks, neurons or parts of neurons
- G06N3/063—Physical realisation, i.e. hardware implementation of neural networks, neurons or parts of neurons using electronic means
- G06N3/0635—Physical realisation, i.e. hardware implementation of neural networks, neurons or parts of neurons using electronic means using analogue means
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING; COUNTING
- G06N—COMPUTER SYSTEMS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N5/00—Computer systems utilising knowledge based models
- G06N5/04—Inference methods or devices
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING; COUNTING
- G06N—COMPUTER SYSTEMS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N5/00—Computer systems utilising knowledge based models
- G06N5/02—Knowledge representation
- G06N5/022—Knowledge engineering, knowledge acquisition
- G06N5/025—Extracting rules from data
-
- G—PHYSICS
- G05—CONTROLLING; REGULATING
- G05B—CONTROL OR REGULATING SYSTEMS IN GENERAL; FUNCTIONAL ELEMENTS OF SUCH SYSTEMS; MONITORING OR TESTING ARRANGEMENTS FOR SUCH SYSTEMS OR ELEMENTS
- G05B13/00—Adaptive control systems, i.e. systems automatically adjusting themselves to have a performance which is optimum according to some preassigned criterion
- G05B13/02—Adaptive control systems, i.e. systems automatically adjusting themselves to have a performance which is optimum according to some preassigned criterion electric
- G05B13/0265—Adaptive control systems, i.e. systems automatically adjusting themselves to have a performance which is optimum according to some preassigned criterion electric the criterion being a learning criterion
- G05B13/027—Adaptive control systems, i.e. systems automatically adjusting themselves to have a performance which is optimum according to some preassigned criterion electric the criterion being a learning criterion using neural networks only
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING; COUNTING
- G06N—COMPUTER SYSTEMS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computer systems based on biological models
- G06N3/12—Computer systems based on biological models using genetic models
- G06N3/126—Genetic algorithms, i.e. information processing using digital simulations of the genetic system
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING; COUNTING
- G06N—COMPUTER SYSTEMS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N7/00—Computer systems based on specific mathematical models
- G06N7/005—Probabilistic networks
-
- G—PHYSICS
- G05—CONTROLLING; REGULATING
- G05B—CONTROL OR REGULATING SYSTEMS IN GENERAL; FUNCTIONAL ELEMENTS OF SUCH SYSTEMS; MONITORING OR TESTING ARRANGEMENTS FOR SUCH SYSTEMS OR ELEMENTS
- G05B13/00—Adaptive control systems, i.e. systems automatically adjusting themselves to have a performance which is optimum according to some preassigned criterion
- G05B13/02—Adaptive control systems, i.e. systems automatically adjusting themselves to have a performance which is optimum according to some preassigned criterion electric
- G05B13/04—Adaptive control systems, i.e. systems automatically adjusting themselves to have a performance which is optimum according to some preassigned criterion electric involving the use of models or simulators
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING; COUNTING
- G06K—RECOGNITION OF DATA; PRESENTATION OF DATA; RECORD CARRIERS; HANDLING RECORD CARRIERS
- G06K9/00—Methods or arrangements for reading or recognising printed or written characters or for recognising patterns, e.g. fingerprints
- G06K9/62—Methods or arrangements for recognition using electronic means
- G06K9/6217—Design or setup of recognition systems and techniques; Extraction of features in feature space; Clustering techniques; Blind source separation
- G06K9/6232—Extracting features by transforming the feature space, e.g. multidimensional scaling; Mappings, e.g. subspace methods
- G06K9/6251—Extracting features by transforming the feature space, e.g. multidimensional scaling; Mappings, e.g. subspace methods based on a criterion of topology preservation, e.g. multidimensional scaling, self-organising maps
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| Zhang et al. | Safe reinforcement learning with stability guarantee for motion planning of autonomous vehicles | |
| US11086938B2 (en) | Interpreting human-robot instructions | |
| Finn et al. | Generalizing skills with semi-supervised reinforcement learning | |
| Thanasutives et al. | Adversarial multi-task learning enhanced physics-informed neural networks for solving partial differential equations | |
| US10481565B2 (en) | Methods and systems for nonlinear adaptive control and filtering | |
| Higuera et al. | Synthesizing neural network controllers with probabilistic model-based reinforcement learning | |
| US12360529B2 (en) | Predictive modeling of aircraft dynamics | |
| Rybkin et al. | Learning what you can do before doing anything | |
| Bacanin et al. | Convolutional neural networks hyperparameters optimization using sine cosine algorithm | |
| Bashir et al. | An overview of inverse reinforcement learning techniques | |
| Alexander et al. | A Comprehensive Survey of Path Planning Algorithms for Autonomous Systems and Mobile Robots: Traditional and Modern Approaches | |
| Kwon et al. | Brain-inspired hyperdimensional computing in the wild: Lightweight symbolic learning for sensorimotor controls of wheeled robots | |
| Hu et al. | Incremental learning framework for autonomous robots based on Q-learning and the adaptive kernel linear model | |
| Yang et al. | Mpr-rl: Multi-prior regularized reinforcement learning for knowledge transfer | |
| Röfer et al. | Bayesian optimization for sample-efficient policy improvement in robotic manipulation | |
| Mittal et al. | Meta-Learning for Few-Shot Adaptation in Robotic Control Tasks | |
| Bar | Self-improving agentic AI through Bayesian meta-learning | |
| Bashir et al. | Inverse reinforcement learning through max-margin algorithm | |
| Yusof et al. | Formulation of a lightweight hybrid AI algorithm towards self-learning autonomous systems | |
| Yu et al. | Deep Q‐Network with Predictive State Models in Partially Observable Domains | |
| Hoang et al. | Deep reinforcement learning and its applications | |
| US20240119363A1 (en) | System and process for deconfounded imitation learning | |
| Fernández et al. | Noise-based reward-modulated learning | |
| Gao et al. | Integrating GAT and LSTM for Robotic Arms Motion Planning | |
| Sabeeh | Enhancing Robotic Grasping Performance through Data-Driven Analysis |