es
Feedback
Data science research papers

Data science research papers

Ir al canal en Telegram

Machine learning and data science research papers Key ML and AI papers with code and GitHub repos. Simple way to follow current research. Join 👉 https://rebrand.ly/bigdatachannels DMCA: @disclosure_bds Contact: @mldatascientist

Mostrar más
3 125
Suscriptores
+424 horas
+217 días
+11230 días
Atraer Suscriptores
agosto '26
agosto '26
+158
en 0 canales
julio '26
+149
en 0 canales
Get PRO
junio '26
+97
en 0 canales
Get PRO
mayo '26
+104
en 1 canales
Get PRO
abril '26
+114
en 0 canales
Get PRO
marzo '26
+101
en 0 canales
Get PRO
febrero '26
+82
en 0 canales
Get PRO
enero '26
+118
en 9 canales
Get PRO
diciembre '25
+115
en 0 canales
Get PRO
noviembre '25
+112
en 0 canales
Get PRO
octubre '25
+43
en 0 canales
Get PRO
septiembre '25
+7
en 0 canales
Get PRO
agosto '25
+3
en 0 canales
Get PRO
julio '25
+3
en 0 canales
Get PRO
junio '25
+2
en 0 canales
Get PRO
mayo '25
+3
en 0 canales
Get PRO
abril '25
+15
en 0 canales
Get PRO
marzo '25
+112
en 0 canales
Get PRO
febrero '25
+153
en 0 canales
Get PRO
enero '25
+187
en 0 canales
Get PRO
diciembre '24
+179
en 0 canales
Get PRO
noviembre '24
+165
en 0 canales
Get PRO
octubre '24
+136
en 0 canales
Get PRO
septiembre '24
+108
en 0 canales
Get PRO
agosto '24
+114
en 0 canales
Get PRO
julio '24
+139
en 0 canales
Get PRO
junio '24
+115
en 0 canales
Get PRO
mayo '24
+132
en 1 canales
Get PRO
abril '24
+109
en 0 canales
Get PRO
marzo '24
+146
en 0 canales
Get PRO
febrero '24
+183
en 0 canales
Get PRO
enero '24
+228
en 0 canales
Get PRO
diciembre '23
+171
en 1 canales
Get PRO
noviembre '23
+28
en 0 canales
Get PRO
octubre '23
+28
en 0 canales
Get PRO
septiembre '23
+504
en 0 canales
Fecha
Crecimiento de Suscriptores
Menciones
Canales
30 agosto0
29 agosto+5
28 agosto+10
27 agosto+4
26 agosto+2
25 agosto+4
24 agosto+1
23 agosto+4
22 agosto+2
21 agosto+3
20 agosto+8
19 agosto+3
18 agosto+3
17 agosto+9
16 agosto+1
15 agosto+4
14 agosto+6
13 agosto+7
12 agosto+5
11 agosto+6
10 agosto+2
09 agosto+4
08 agosto+6
07 agosto+32
06 agosto+5
05 agosto+5
04 agosto+7
03 agosto+2
02 agosto+3
01 agosto+5
Publicaciones del Canal
Does VLA Even Know the Basics? Measuring Commonsense and World Knowledge Retention in Vision-Language-Action Models 📅 Public
Does VLA Even Know the Basics? Measuring Commonsense and World Knowledge Retention in Vision-Language-Action Models 📅 Publication Date: Jun 17, 2026 📑 Paper: https://arxiv.org/pdf/2606.19297.pdf 🔗 Code: N/A 📝 Description: Act2Answer protocol evaluates embodied vision-language-action models by having agents answer questions through physical actions, revealing knowledge retention and generalization patterns across different semantic categories.

2
Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent 📅 Publication Date: Jun 29
Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent 📅 Publication Date: Jun 29, 2026 📑 Paper: https://arxiv.org/pdf/2606.30616.pdf 🔗 Code: N/A 📝 Description: Agents-A1, a 35B Mixture-of-Experts Agentic Model, achieves trillion-parameter-level performance through long-horizon trajectory scaling and heterogeneous agent ability scaling via a three-stage training approach involving supervised fine-tuning, domain-level teacher models, and multi-teacher distil...
95
3
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation 📅 Publication Date: Jun 26, 2026 📑 Paper: https:
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation 📅 Publication Date: Jun 26, 2026 📑 Paper: https://arxiv.org/pdf/2606.28128.pdf 🔗 Code: https://github.com/huggingface 📝 Description: PhysisForcing enhances embodied video generation by enforcing physical consistency through pixel-level trajectory alignment and semantic-level relational alignment losses in a DiT-based framework.
168
4
LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing 📅 Publication Date: Jun 25, 2026 📑 Paper: https://arxiv
LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing 📅 Publication Date: Jun 25, 2026 📑 Paper: https://arxiv.org/pdf/2606.26740.pdf 🔗 Code: N/A 📝 Description: A novel streaming video editing framework enables causal, frame-by-frame editing with stable long-horizon preservation and real-time responsiveness through a three-stage distillation pipeline and AR-oriented mask cache.
199
5
BlockPilot: Instance-Adaptive Policy Learning for Diffusion-based Speculative Decoding 📅 Publication Date: Jun 30, 2026 📑 P
BlockPilot: Instance-Adaptive Policy Learning for Diffusion-based Speculative Decoding 📅 Publication Date: Jun 30, 2026 📑 Paper: https://arxiv.org/pdf/2606.31315.pdf 🔗 Code: N/A 📝 Description: Speculative decoding with adaptive block size selection improves inference efficiency by predicting optimal block sizes from prefilling representations, achieving significant speedup with minimal overhead.
249
6
DOPD: Dual On-policy Distillation 📅 Publication Date: Jun 29, 2026 📑 Paper: https://arxiv.org/pdf/2606.30626.pdf 🔗 Code: h
DOPD: Dual On-policy Distillation 📅 Publication Date: Jun 29, 2026 📑 Paper: https://arxiv.org/pdf/2606.30626.pdf 🔗 Code: https://github.com/huggingface 📝 Description: DOPD addresses privilege illusion in on-policy distillation by dynamically routing token-level supervision between teacher and student policies based on advantage gaps and probabilities, improving capability transfer in large and vision-language models.
290
7
Dockerless: Environment-Free Program Verifier for Coding Agents 📅 Publication Date: Jun 26, 2026 📑 Paper: https://arxiv.org
Dockerless: Environment-Free Program Verifier for Coding Agents 📅 Publication Date: Jun 26, 2026 📑 Paper: https://arxiv.org/pdf/2606.28436.pdf 🔗 Code: https://github.com/huggingface 📝 Description: A Dockerless environment-free agentic patch verifier improves code patch evaluation accuracy and enables effective post-training without execution-based verification costs.
324
8
Agentic Abstention: Do Agents Know When to Stop Instead of Act? 📅 Publication Date: Jun 27, 2026 📑 Paper: https://arxiv.org
Agentic Abstention: Do Agents Know When to Stop Instead of Act? 📅 Publication Date: Jun 27, 2026 📑 Paper: https://arxiv.org/pdf/2606.28733.pdf 🔗 Code: https://github.com/huggingface 📝 Description: Agentic abstention involves determining when an AI agent should cease interaction under uncertainty, requiring sequential decision-making across multiple environments and task types.
355
9
Orca: The World is in Your Mind 📅 Publication Date: Jun 29, 2026 📑 Paper: https://arxiv.org/pdf/2606.30534.pdf 🔗 Code: htt
Orca: The World is in Your Mind 📅 Publication Date: Jun 29, 2026 📑 Paper: https://arxiv.org/pdf/2606.30534.pdf 🔗 Code: https://github.com/huggingface 📝 Description: Orca establishes a unified world latent space through next-state-prediction modeling using multimodal data and demonstrates superior performance in downstream tasks compared to specialized baselines.
361
10
Grouped Query Experts: Mixture-of-Experts on GQA Self-Attention 📅 Publication Date: Jun 18, 2026 📑 Paper: https://arxiv.org
Grouped Query Experts: Mixture-of-Experts on GQA Self-Attention 📅 Publication Date: Jun 18, 2026 📑 Paper: https://arxiv.org/pdf/2606.20945.pdf 🔗 Code: https://github.com/huggingface 📝 Description: Grouped Query Experts (GQE) improves Transformer efficiency by selectively activating query heads based on token content while maintaining key-value cache benefits of grouped-query attention.
377
11
PlanBench-XL: Evaluating Long-Horizon Planning of LLM Tool-Use Agents in Large-Scale Tool Ecosystems 📅 Publication Date: Jun
PlanBench-XL: Evaluating Long-Horizon Planning of LLM Tool-Use Agents in Large-Scale Tool Ecosystems 📅 Publication Date: Jun 21, 2026 📑 Paper: https://arxiv.org/pdf/2606.22388.pdf 🔗 Code: https://github.com/huggingface 📝 Description: PlanBench-XL evaluates large language model agents' ability to plan and adapt in complex tool-rich environments with limited visibility and dynamic disruptions.
392
12
Are We Ready For An Agent-Native Memory System? 📅 Publication Date: Jun 23, 2026 📑 Paper: https://arxiv.org/pdf/2606.24775.
Are We Ready For An Agent-Native Memory System? 📅 Publication Date: Jun 23, 2026 📑 Paper: https://arxiv.org/pdf/2606.24775.pdf 🔗 Code: https://github.com/huggingface 📝 Description: Large language model agents' memory systems have evolved into complex data management frameworks requiring systematic evaluation across multiple modules and workloads to understand their performance characteristics and trade-offs.
395
13
EnterpriseClawBench: Benchmarking Agents from Real Workplace Sessions 📅 Publication Date: Jun 22, 2026 📑 Paper: https://arx
EnterpriseClawBench: Benchmarking Agents from Real Workplace Sessions 📅 Publication Date: Jun 22, 2026 📑 Paper: https://arxiv.org/pdf/2606.23654.pdf 🔗 Code: https://github.com/huggingface 📝 Description: EnterpriseClawBench presents a benchmark for enterprise agents based on real-world sessions with 852 reproducible tasks, emphasizing comprehensive evaluation metrics beyond single performance scores.
376
14
OpenRath: Session-Centered Runtime State for Agent Systems 📅 Publication Date: Jun 17, 2026 📑 Paper: https://arxiv.org/pdf/
OpenRath: Session-Centered Runtime State for Agent Systems 📅 Publication Date: Jun 17, 2026 📑 Paper: https://arxiv.org/pdf/2606.19409.pdf 🔗 Code: https://github.com/huggingface 📝 Description: OpenRath introduces a PyTorch-like programming model for multi-agent systems using Session as a central runtime abstraction that enables explicit fork, merge, and replay operations while recording comprehensive execution state.
390
15
MemSlides: A Hierarchical Memory Driven Agent Framework for Personalized Slide Generation with Multi-turn Local Revision 📅 P
MemSlides: A Hierarchical Memory Driven Agent Framework for Personalized Slide Generation with Multi-turn Local Revision 📅 Publication Date: Jun 15, 2026 📑 Paper: https://arxiv.org/pdf/2606.17162.pdf 🔗 Code: https://github.com/huggingface 📝 Description: MemSlides presents a hierarchical memory framework for personalized presentation agents that separates long-term user profiles, working memory for session constraints, and tool memory for reusable execution experiences to enable stable personalization and reliable local edits.
436
16
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs 📅 Publication Date: May 7, 2026 📑 Paper: https:/
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs 📅 Publication Date: May 7, 2026 📑 Paper: https://arxiv.org/pdf/2606.27378.pdf 🔗 Code: N/A 📝 Description: An axiomatic evaluation framework reveals systematic failures in latent thought representations of LLMs across multiple reasoning tasks, demonstrating that current representations fail to satisfy fundamental functional axioms consistently across different model architectures.
395
17
EnterpriseClawBench: Benchmarking Agents from Real Workplace Sessions 📅 Publication Date: Jun 22, 2026 📑 Paper: https://arxiv.org/pdf/2606.23654.pdf 🔗 Code: https://github.com/huggingface 📝 Description: EnterpriseClawBench presents a benchmark for enterprise agents based on real-world sessions with 852 reproducible tasks, emphasizing comprehensive evaluation metrics beyond single performance scores.
418
18
Heterogeneous Scientific Foundation Model Collaboration 📅 Publication Date: Apr 30, 2026 📑 Paper: https://arxiv.org/pdf/260
Heterogeneous Scientific Foundation Model Collaboration 📅 Publication Date: Apr 30, 2026 📑 Paper: https://arxiv.org/pdf/2604.27351.pdf 🔗 Code: https://github.com/Violet24K/Eywa 📝 Description: Eywa is a heterogeneous agentic framework that extends language-centric systems to scientific foundation models by integrating domain-specific models with language-based reasoning interfaces for improved performance across diverse scientific domains.
444
19
OpenSearch-VL: An Open Recipe for Frontier Multimodal Search Agents 📅 Publication Date: May 6, 2026 📑 Paper: https://arxiv.org/pdf/2605.05185.pdf 🔗 Code: https://github.com/shawn0728/OpenSearch-VL 📝 Description: OpenSearch-VL presents an open-source framework for training advanced multimodal search agents using reinforcement learning, featuring specialized data curation, diverse tool environments, and a novel training algorithm that improves performance across multiple benchmarks.
427
20
WorldOlympiad: Can Your World Model Survive a Triathlon? 📅 Publication Date: Jun 9, 2026 📑 Paper: https://arxiv.org/pdf/260
WorldOlympiad: Can Your World Model Survive a Triathlon? 📅 Publication Date: Jun 9, 2026 📑 Paper: https://arxiv.org/pdf/2606.11129 💻 Project Page: https://alibaba-damo-academy.github.io/WorldOlympiad/ 📝 Description: The paper introduces WorldOlympiad, a comprehensive benchmark for evaluating video-based world models. The problem with current generative models is that they often focus on visual quality, but lack physical faithfulness, geometric consistency, and interaction fidelity. To address this gap, WorldOlympiad decomposes world-model evaluation into three dimensions: physical faithfulness, geometric consistency, and interaction fidelity. WorldOlympiad covers three major downstream scenarios, including gaming, robotics, and general real-world videos, capturing diverse challenges from interactive control and embodied manipulation to open-domain motion and camera dynamics. #WorldModelEvaluation #VideoBasedWorldModels #PhysicalFaithfulness #GeometricConsistency
421