Models
Datasets
Spaces
Docs
Enterprise
Pricing
Log In
Sign Up

Collections

Discover the best community collections!

Collections including paper arxiv:2510.17269

cabinet-data_curation

Skywork-Reward-V2: Scaling Preference Data Curation via Human-AI Synergy

Paper • 2507.01352 • Published Jul 2, 2025 • 57
A Data-Centric Framework for Addressing Phonetic and Prosodic Challenges in Russian Speech Generative Models

Paper • 2507.13563 • Published Jul 17, 2025 • 53
Scaling Laws for Optimal Data Mixtures

Paper • 2507.09404 • Published Jul 12, 2025 • 37
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation

Paper • 2511.14993 • Published Nov 19, 2025 • 231

Easy Dataset: A Unified and Extensible Framework for Synthesizing LLM Fine-Tuning Data from Unstructured Documents

Paper • 2507.04009 • Published Jul 5, 2025 • 54
FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

Read Later Stack

Demystifying Reinforcement Learning in Agentic Reasoning

Paper • 2510.11701 • Published Oct 13, 2025 • 33
Self-Improving LLM Agents at Test-Time

Paper • 2510.07841 • Published Oct 9, 2025 • 10
Making Mathematical Reasoning Adaptive

Paper • 2510.04617 • Published Oct 6, 2025 • 23
DocReward: A Document Reward Model for Structuring and Stylizing

Paper • 2510.11391 • Published Oct 13, 2025 • 27

TempoFunk/hdvila-100M

Viewer • Updated Dec 2, 2023 • 30.1M • 209 • 17
HuggingFaceM4/FineVision

Viewer • Updated Oct 21, 2025 • 24.2M • 107k • 470
kakaobrain/coyo-700m

Viewer • Updated Aug 30, 2022 • 747M • 1.64k • 155
FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

FineVision dataset

FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

about 21 hours ago

LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios

Paper • 2509.09926 • Published Sep 12, 2025 • 14
What Breaks Knowledge Graph based RAG? Empirical Insights into Reasoning under Incomplete Knowledge

Paper • 2508.08344 • Published Aug 11, 2025
MemMamba: Rethinking Memory Patterns in State Space Model

Paper • 2510.03279 • Published Sep 28, 2025 • 73
When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs

Paper • 2510.07499 • Published Oct 8, 2025 • 48

Data and other things

MegaPairs: Massive Data Synthesis For Universal Multimodal Retrieval

Paper • 2412.14475 • Published Dec 19, 2024 • 57
How to Synthesize Text Data without Model Collapse?

Paper • 2412.14689 • Published Dec 19, 2024 • 53
Token-Budget-Aware LLM Reasoning

Paper • 2412.18547 • Published Dec 24, 2024 • 46
WavePulse: Real-time Content Analytics of Radio Livestreams

Paper • 2412.17998 • Published Dec 23, 2024 • 11

cabinet-data_curation

Skywork-Reward-V2: Scaling Preference Data Curation via Human-AI Synergy

Paper • 2507.01352 • Published Jul 2, 2025 • 57
A Data-Centric Framework for Addressing Phonetic and Prosodic Challenges in Russian Speech Generative Models

Paper • 2507.13563 • Published Jul 17, 2025 • 53
Scaling Laws for Optimal Data Mixtures

Paper • 2507.09404 • Published Jul 12, 2025 • 37
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation

Paper • 2511.14993 • Published Nov 19, 2025 • 231

FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

Easy Dataset: A Unified and Extensible Framework for Synthesizing LLM Fine-Tuning Data from Unstructured Documents

Paper • 2507.04009 • Published Jul 5, 2025 • 54
FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

FineVision dataset

FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

Read Later Stack

Demystifying Reinforcement Learning in Agentic Reasoning

Paper • 2510.11701 • Published Oct 13, 2025 • 33
Self-Improving LLM Agents at Test-Time

Paper • 2510.07841 • Published Oct 9, 2025 • 10
Making Mathematical Reasoning Adaptive

Paper • 2510.04617 • Published Oct 6, 2025 • 23
DocReward: A Document Reward Model for Structuring and Stylizing

Paper • 2510.11391 • Published Oct 13, 2025 • 27

about 21 hours ago

LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios

Paper • 2509.09926 • Published Sep 12, 2025 • 14
What Breaks Knowledge Graph based RAG? Empirical Insights into Reasoning under Incomplete Knowledge

Paper • 2508.08344 • Published Aug 11, 2025
MemMamba: Rethinking Memory Patterns in State Space Model

Paper • 2510.03279 • Published Sep 28, 2025 • 73
When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs

Paper • 2510.07499 • Published Oct 8, 2025 • 48

TempoFunk/hdvila-100M

Viewer • Updated Dec 2, 2023 • 30.1M • 209 • 17
HuggingFaceM4/FineVision

Viewer • Updated Oct 21, 2025 • 24.2M • 107k • 470
kakaobrain/coyo-700m

Viewer • Updated Aug 30, 2022 • 747M • 1.64k • 155
FineVision: Open Data Is All You Need

Paper • 2510.17269 • Published Oct 20, 2025 • 75

Data and other things

MegaPairs: Massive Data Synthesis For Universal Multimodal Retrieval

Paper • 2412.14475 • Published Dec 19, 2024 • 57
How to Synthesize Text Data without Model Collapse?

Paper • 2412.14689 • Published Dec 19, 2024 • 53
Token-Budget-Aware LLM Reasoning

Paper • 2412.18547 • Published Dec 24, 2024 • 46
WavePulse: Real-time Content Analytics of Radio Livestreams

Paper • 2412.17998 • Published Dec 23, 2024 • 11

Previous
1
2
Next

Company

TOS Privacy About Careers

Website

Models Datasets Spaces Pricing Docs