Picture for Mikhail Konenkov

Mikhail Konenkov

ORCESTRA: VLM-driven Visual Robot programming in Mixed Reality

Add code
Aug 01, 2026
Viaarxiv icon

AgenticFocus: Object-Preserving Mixed Reality Synthesis from Human FPV Video for Dexterous Humanoid Learning

Add code
Jul 13, 2026
Viaarxiv icon

Closed-Loop Verbal Reinforcement Learning for Task-Level Robotic Planning

Add code
Mar 23, 2026
Viaarxiv icon

PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models

Add code
Sep 17, 2025
Figure 1 for PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
Figure 2 for PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
Figure 3 for PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
Figure 4 for PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
Viaarxiv icon

Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots

Add code
Sep 16, 2024
Figure 1 for Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots
Figure 2 for Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots
Figure 3 for Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots
Figure 4 for Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots
Viaarxiv icon

VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications

Add code
May 19, 2024
Viaarxiv icon

Co-driver: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes

Add code
May 09, 2024
Figure 1 for Co-driver: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
Figure 2 for Co-driver: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
Figure 3 for Co-driver: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
Figure 4 for Co-driver: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
Viaarxiv icon

FlockGPT: Guiding UAV Flocking with Linguistic Orchestration

Add code
May 09, 2024
Figure 1 for FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
Figure 2 for FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
Figure 3 for FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
Figure 4 for FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
Viaarxiv icon

HawkDrive: A Transformer-driven Visual Perception System for Autonomous Driving in Night Scene

Add code
Apr 06, 2024
Figure 1 for HawkDrive: A Transformer-driven Visual Perception System for Autonomous Driving in Night Scene
Figure 2 for HawkDrive: A Transformer-driven Visual Perception System for Autonomous Driving in Night Scene
Figure 3 for HawkDrive: A Transformer-driven Visual Perception System for Autonomous Driving in Night Scene
Figure 4 for HawkDrive: A Transformer-driven Visual Perception System for Autonomous Driving in Night Scene
Viaarxiv icon

CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI

Add code
Jan 29, 2024
Figure 1 for CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI
Figure 2 for CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI
Figure 3 for CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI
Figure 4 for CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI
Viaarxiv icon