Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Sidharth Ramesh

MedPAO: A Protocol-Driven Agent for Structuring Medical Reports

Oct 06, 2025

Shrish Shrinath Vaidya, Gowthamaan Palani, Sidharth Ramesh, Velmurugan Balasubramanian, Minmini Selvam, Gokulraja Srinivasaraja, Ganapathy Krishnamurthi

Abstract:The deployment of Large Language Models (LLMs) for structuring clinical data is critically hindered by their tendency to hallucinate facts and their inability to follow domain-specific rules. To address this, we introduce MedPAO, a novel agentic framework that ensures accuracy and verifiable reasoning by grounding its operation in established clinical protocols such as the ABCDEF protocol for CXR analysis. MedPAO decomposes the report structuring task into a transparent process managed by a Plan-Act-Observe (PAO) loop and specialized tools. This protocol-driven method provides a verifiable alternative to opaque, monolithic models. The efficacy of our approach is demonstrated through rigorous evaluation: MedPAO achieves an F1-score of 0.96 on the critical sub-task of concept categorization. Notably, expert radiologists and clinicians rated the final structured outputs with an average score of 4.52 out of 5, indicating a level of reliability that surpasses baseline approaches relying solely on LLM-based foundation models. The code is available at: https://github.com/MiRL-IITM/medpao-agent

* Lecture Notes in Computer Science, vol 16147, 2025. Springer, Cham
* Paper published at "Agentic AI for Medicine" Workshop, MICCAI 2025

Via

Access Paper or Ask Questions

Autonomous Control of a Novel Closed Chain Five Bar Active Suspension via Deep Reinforcement Learning

Jun 27, 2024

Nishesh Singh, Sidharth Ramesh, Abhishek Shankar, Jyotishka Duttagupta, Leander Stephen D'Souza, Sanjay Singh

Figure 1 for Autonomous Control of a Novel Closed Chain Five Bar Active Suspension via Deep Reinforcement Learning

Figure 2 for Autonomous Control of a Novel Closed Chain Five Bar Active Suspension via Deep Reinforcement Learning

Figure 3 for Autonomous Control of a Novel Closed Chain Five Bar Active Suspension via Deep Reinforcement Learning

Figure 4 for Autonomous Control of a Novel Closed Chain Five Bar Active Suspension via Deep Reinforcement Learning

Abstract:Planetary exploration requires traversal in environments with rugged terrains. In addition, Mars rovers and other planetary exploration robots often carry sensitive scientific experiments and components onboard, which must be protected from mechanical harm. This paper deals with an active suspension system focused on chassis stabilisation and an efficient traversal method while encountering unavoidable obstacles. Soft Actor-Critic (SAC) was applied along with Proportional Integral Derivative (PID) control to stabilise the chassis and traverse large obstacles at low speeds. The model uses the rover's distance from surrounding obstacles, the height of the obstacle, and the chassis' orientation to actuate the control links of the suspension accurately. Simulations carried out in the Gazebo environment are used to validate the proposed active system.

* 15 pages, 11 figures

Via

Access Paper or Ask Questions

Fast and Accurate Camera Scene Detection on Smartphones

May 17, 2021

Angeline Pouget, Sidharth Ramesh, Maximilian Giang, Ramithan Chandrapalan, Toni Tanner, Moritz Prussing, Radu Timofte, Andrey Ignatov

Figure 1 for Fast and Accurate Camera Scene Detection on Smartphones

Figure 2 for Fast and Accurate Camera Scene Detection on Smartphones

Figure 3 for Fast and Accurate Camera Scene Detection on Smartphones

Figure 4 for Fast and Accurate Camera Scene Detection on Smartphones

Abstract:AI-powered automatic camera scene detection mode is nowadays available in nearly any modern smartphone, though the problem of accurate scene prediction has not yet been addressed by the research community. This paper for the first time carefully defines this problem and proposes a novel Camera Scene Detection Dataset (CamSDD) containing more than 11K manually crawled images belonging to 30 different scene categories. We propose an efficient and NPU-friendly CNN model for this task that demonstrates a top-3 accuracy of 99.5% on this dataset and achieves more than 200 FPS on the recent mobile SoCs. An additional in-the-wild evaluation of the obtained solution is performed to analyze its performance and limitation in the real-world scenarios. The dataset and pre-trained models used in this paper are available on the project website.

Via

Access Paper or Ask Questions