Customer-obsessed science
Research areas
-
August 26, 20265 min readDiscounting the opinions of LLM judges with highly correlated outputs ensures that panels of judges reflect a true diversity of perspectives.
-
August 21, 20269 min read
-
July 30, 20268 min read
-
-
July 9, 202610 min read
Featured news
-
IEEE CASE 20262026While many existing grasping models can be highly reliable in picking objects in most cases, challenging scenarios persist in industrial automation where objects are difficult to grasp—such as when positioned in corners, occluded by other items, or tightly clustered. These challenges are prevalent in smart manufacturing and logistics systems, where current robotic systems often require costly human intervention
-
ICML 2026 Workshop on Weight-Space Symmetries2026Multi-task model merging combines separately trained expert models into a single model that handles all tasks without co-training. Standard practice merges experts at their optimal validation loss. We challenge this convention by systematically studying how training duration of domain experts affects the quality of the merged model. We fine-tune experts on five domains (Math, Code, Instruction Following
-
ACL 2026 Workshop on Sustainable and Efficient Language, Vision, and Action Models (SELVA)2026Optimizing Large Language Models (LLMs) for production AI agent deployment demands substantial computational resources and specialized human expertise (e.g., prompt engineering). Self-evolution offers a promising solution by enabling agents to autonomously enhance capabilities through structured feedback, improving performance without expensive manual optimization. However, most existing self-evolving agents
-
2026Understanding the behavior and logical structure of complex algorithms is a fundamental challenge in industrial systems. Recent advancements in large language models (LLMs) have demonstrated remarkable code understanding capabilities. However, their potential for reverse engineering algorithms into interpretable causal structures remains unexplored. In this work, we develop a multi-agent framework, RECoRD
-
Transactions on Machine Learning Research2026Large Reasoning Models (LRMs) excel at complex reasoning tasks, but their efficiency is often hampered by overly verbose outputs. Prior steering methods attempt to address this issue by applying a single, global vector to hidden representations—an approach grounded in the restrictive linear representation hypothesis. In this work, we introduce FlowSteer, a nonlinear steering method that goes beyond uniform
Collaborations
View allWhether you're a faculty member or student, there are number of ways you can engage with Amazon.
View all