-
TradingAgents: Multi-Agents LLM Financial Trading Framework
Paper • 2412.20138 • Published • 122 -
OpenDevin: An Open Platform for AI Software Developers as Generalist Agents
Paper • 2407.16741 • Published • 84 -
Kronos: A Foundation Model for the Language of Financial Markets
Paper • 2508.02739 • Published • 52 -
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
Paper • 2607.19191 • Published • 311
Collections
Discover the best community collections!
Collections including paper arxiv:2607.19191
-
AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report
Paper • 2607.18367 • Published • 60 -
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
Paper • 2607.19191 • Published • 311 -
Vidu S1: A Real-Time Interactive Video Generation Model
Paper • 2607.03118 • Published • 144 -
AlayaWorld: Long-Horizon and Playable Video World Generation
Paper • 2607.06291 • Published • 92
-
Code as Agent Harness
Paper • 2605.18747 • Published • 226 -
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 195 -
From Context to Skills: Can Language Models Learn from Context Skillfully?
Paper • 2604.27660 • Published • 171 -
PhysBrain 1.0 Technical Report
Paper • 2605.15298 • Published • 145
-
LLM Pruning and Distillation in Practice: The Minitron Approach
Paper • 2408.11796 • Published • 61 -
TableBench: A Comprehensive and Complex Benchmark for Table Question Answering
Paper • 2408.09174 • Published • 53 -
To Code, or Not To Code? Exploring Impact of Code in Pre-training
Paper • 2408.10914 • Published • 45 -
Open-FinLLMs: Open Multimodal Large Language Models for Financial Applications
Paper • 2408.11878 • Published • 64
-
AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis
Paper • 2607.28618 • Published • 302 -
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents
Paper • 2607.28227 • Published • 302 -
Metis: Memory Foundation Model
Paper • 2607.26760 • Published • 271 -
Kimi K3: Open Frontier Intelligence
Paper • 2607.24653 • Published • 491
-
GLM-5: from Vibe Coding to Agentic Engineering
Paper • 2602.15763 • Published • 208 -
zai-org/GLM-5.2
Text Generation • 753B • Updated • 2.69M • • 4.96k -
SkillOpt: Executive Strategy for Self-Evolving Agent Skills
Paper • 2605.23904 • Published • 264 -
SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion
Paper • 2503.11576 • Published • 166
-
FantasyWorld: Geometry-Consistent World Modeling via Unified Video and 3D Prediction
Paper • 2509.21657 • Published • 5 -
FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
Paper • 2504.04842 • Published • 35 -
FantasyHSI: Video-Generation-Centric 4D Human Synthesis In Any Scene through A Graph-based Multi-Agent Framework
Paper • 2509.01232 • Published -
FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation
Paper • 2508.11255 • Published • 11
-
AI for Auto-Research: Roadmap & User Guide
Paper • 2605.18661 • Published • 72 -
StableVLA: Towards Robust Vision-Language-Action Models without Extra Data
Paper • 2605.18287 • Published • 15 -
MixSD: Mixed Contextual Self-Distillation for Knowledge Injection
Paper • 2605.16865 • Published • 10 -
MolmoPoint: Better Pointing for VLMs with Grounding Tokens
Paper • 2603.28069 • Published • 9
-
TradingAgents: Multi-Agents LLM Financial Trading Framework
Paper • 2412.20138 • Published • 122 -
OpenDevin: An Open Platform for AI Software Developers as Generalist Agents
Paper • 2407.16741 • Published • 84 -
Kronos: A Foundation Model for the Language of Financial Markets
Paper • 2508.02739 • Published • 52 -
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
Paper • 2607.19191 • Published • 311
-
AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis
Paper • 2607.28618 • Published • 302 -
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents
Paper • 2607.28227 • Published • 302 -
Metis: Memory Foundation Model
Paper • 2607.26760 • Published • 271 -
Kimi K3: Open Frontier Intelligence
Paper • 2607.24653 • Published • 491
-
AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report
Paper • 2607.18367 • Published • 60 -
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
Paper • 2607.19191 • Published • 311 -
Vidu S1: A Real-Time Interactive Video Generation Model
Paper • 2607.03118 • Published • 144 -
AlayaWorld: Long-Horizon and Playable Video World Generation
Paper • 2607.06291 • Published • 92
-
GLM-5: from Vibe Coding to Agentic Engineering
Paper • 2602.15763 • Published • 208 -
zai-org/GLM-5.2
Text Generation • 753B • Updated • 2.69M • • 4.96k -
SkillOpt: Executive Strategy for Self-Evolving Agent Skills
Paper • 2605.23904 • Published • 264 -
SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion
Paper • 2503.11576 • Published • 166
-
Code as Agent Harness
Paper • 2605.18747 • Published • 226 -
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 195 -
From Context to Skills: Can Language Models Learn from Context Skillfully?
Paper • 2604.27660 • Published • 171 -
PhysBrain 1.0 Technical Report
Paper • 2605.15298 • Published • 145
-
FantasyWorld: Geometry-Consistent World Modeling via Unified Video and 3D Prediction
Paper • 2509.21657 • Published • 5 -
FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
Paper • 2504.04842 • Published • 35 -
FantasyHSI: Video-Generation-Centric 4D Human Synthesis In Any Scene through A Graph-based Multi-Agent Framework
Paper • 2509.01232 • Published -
FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation
Paper • 2508.11255 • Published • 11
-
LLM Pruning and Distillation in Practice: The Minitron Approach
Paper • 2408.11796 • Published • 61 -
TableBench: A Comprehensive and Complex Benchmark for Table Question Answering
Paper • 2408.09174 • Published • 53 -
To Code, or Not To Code? Exploring Impact of Code in Pre-training
Paper • 2408.10914 • Published • 45 -
Open-FinLLMs: Open Multimodal Large Language Models for Financial Applications
Paper • 2408.11878 • Published • 64
-
AI for Auto-Research: Roadmap & User Guide
Paper • 2605.18661 • Published • 72 -
StableVLA: Towards Robust Vision-Language-Action Models without Extra Data
Paper • 2605.18287 • Published • 15 -
MixSD: Mixed Contextual Self-Distillation for Knowledge Injection
Paper • 2605.16865 • Published • 10 -
MolmoPoint: Better Pointing for VLMs with Grounding Tokens
Paper • 2603.28069 • Published • 9