AIBullisharXiv – CS AI · May 117/10
🧠Researchers present Trajectory-Shaped Discrete Flow Matching (TS-DFM), a technique that improves text generation efficiency by using an energy-based guidance system during training to select better token transformation paths. The method enables a compact student model to achieve 32% lower perplexity than a 1,024-step teacher while running 128x faster at just 8 steps, setting new benchmarks for discrete generation tasks.
🏢 Perplexity
AIBullisharXiv – CS AI · Mar 177/10
🧠Researchers introduce AgentDiet, a trajectory reduction technique that cuts computational costs for LLM-based agents by 39.9%-59.7% in input tokens and 21.1%-35.9% in total costs while maintaining performance. The approach removes redundant and expired information from agent execution trajectories during inference time.
AIBullisharXiv – CS AI · Jun 256/10
🧠Researchers introduce ExTra, a reinforcement learning framework that improves language model reasoning by extracting exploration signals from model rollouts. The method combines novelty rewards for diverse solutions with entropy-guided trajectory regeneration, achieving 5-7 point improvements over baseline GRPO across mathematical reasoning benchmarks.
AINeutralarXiv – CS AI · Jun 256/10
🧠Researchers propose RA-QAGC, a quantum-inspired algorithm combining graph condensation with reinforcement learning to optimize UAV trajectory coordination in interference-limited networks. The approach demonstrates 15% throughput gains and 34% improvements in priority-user performance compared to existing methods, addressing scalability challenges in real-time multi-UAV coordination.
AINeutralarXiv – CS AI · Jun 235/10
🧠Researchers introduce Active-Sensing Deferred-Decision Trajectory Optimization (AS-DDTO), an advanced planning algorithm that optimizes mobile sensing system trajectories for target identification while maintaining reachability under resource constraints. The method enhances traditional DDTO by incorporating information-acquisition objectives, enabling earlier target identification through strategic path planning in uncertain sensing environments.
AINeutralarXiv – CS AI · Jun 46/10
🧠Researchers have developed a framework using large language models to automatically translate natural language mission descriptions into executable trajectory optimization code for spacecraft operations. The approach demonstrates high success rates in formulating complex space mission problems, potentially reducing the domain expertise required for trajectory design in autonomous space exploration.
AIBullisharXiv – CS AI · May 116/10
🧠WebClipper is a new framework that optimizes web agent trajectories by pruning redundant reasoning steps through graph-based analysis, reducing tool-call rounds by approximately 20% while maintaining or improving accuracy. The approach models agent search processes as directed acyclic graphs and introduces an F-AE Score metric to measure the balance between accuracy and efficiency in web agent design.
AIBullisharXiv – CS AI · Mar 176/10
🧠Researchers introduce SmoothVLA, a new reinforcement learning framework that improves robot control by optimizing both task performance and motion smoothness. The system addresses the trade-off between stability and exploration in Vision-Language-Action models, achieving 13.8% better smoothness than standard RL methods.
AINeutralarXiv – CS AI · Mar 54/10
🧠Researchers have developed Q-SVMPC, a new Model Predictive Control method that combines reinforcement learning with Stein variational inference to improve trajectory optimization. The approach addresses limitations in existing MPC methods that often converge to single solutions, instead maintaining diverse solution paths for better performance in robotics applications.