Paper Archive

Browse and export your curated research paper collection

336
Archived Days
3323
Total Papers
7.6
Avg Score
12
Categories

Export Archive Data

Download your archived papers in various formats

JSON: Complete data with analysis | CSV: Tabular data for analysis | Markdown: Human-readable reports | BibTeX: Academic citations
Browse by Date

Papers for September 3, 2026

5 papers found

[object Object], [object Object], [object Object], [object Object], [object Object], [object Object], [object Object], [object Object] 9/1/2026 huggingface

natural language processing

Ensuring factuality remains a critical challenge for deploying LLMs in high-stakes settings. Existing hallucination detectors usually operate at a single level: claim-level methods provide interpretable factual units, while span-level methods localize unsupported text. Bridging these views is costly...

Keywords: detection

[object Object], [object Object], [object Object], [object Object], [object Object] 9/1/2026 huggingface

natural language processing

We present H3-World, an efficient framework that turns the 33B MiniMax-H3 video generator into an interactive world model. Our key finding is that, as large video generators become more capable, language is emerging as a natural interface for control. MiniMax-H3, for example, already supports zero-s...

Keywords: attention, pretraining

[object Object], [object Object], [object Object], [object Object] 9/2/2026 huggingface

natural language processing

Ride-sharing, which allows multiple passengers with different origin-destination (OD) pairs to share a single vehicle, is a challenging operational problem, as it requires orders with different OD pairs to be efficiently bundled and assigned to vehicles under uncertain and varying scenarios. Althoug...

Keywords: reinforcement learning

[object Object], [object Object], [object Object], [object Object], [object Object], [object Object], [object Object] 9/1/2026 huggingface

natural language processing

AI tutors are most useful when they adapt to each student's strengths, weaknesses, and preferred guidance, but evidence about which guidance works for which student is sparse, slow, and costly to collect from real learners. Student simulators can provide this signal as a proxy, yet existing approach...

Keywords: gpt, reinforcement learning

[object Object], [object Object], [object Object], [object Object], [object Object], [object Object] 9/2/2026 huggingface

computer vision

Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for large language model (LLM) post-training, but its reliance on coarse outcome rewards leads to limited guidance on intermediate reasoning processes. Existing approaches such as process reward modeling and on-...

Keywords: reinforcement learning
Loading...

Preparing your export...