LLMs

Latest breakthroughs in Large Language Models

Articles

Detecting and Controlling Sycophancy with Cascading Linear Features
Project Auto-World: Towards Automated Benchmarking of Neural Relational Reasoners
The Hitchhiker's Guide to Agentic AI: From Foundations to Systems
Neuro-Symbolic Drive: Rule-Grounded Faithful Reasoning for Driving VLAs
RIFT-Bench: Dynamic Red-teaming for Agentic AI Systems
Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies
Measuring Curriculum Alignment across Topical Coverage, Competency, and Cognitive Depth: A Longitudinal Framework Applied to CS2013 and CS2023
Deontic Policies for Runtime Governance of Agentic AI Systems
CaVe-VLM-CoT: An Interpretable Vision-Language Model Framework with Evidence-Grounded Reasoning
NAVI-Orbital: First In-Orbit Demonstration of a Zero-Shot Vision-Language Model for Autonomous Earth Observation