Running 232 The ultimate guide to RL environments: building and scaling them in the LLM era π 232 Building and scaling RL environments for LLM training
Running on CPU Upgrade 275 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens π 275 Visualize syntheticβdata experiments as an interactive bookshelf
OpenMed/OpenMed-PII-BioClinicalModern-Large-395M-v1 Token Classification β’ 0.4B β’ Updated Jan 13 β’ 18.2k β’ β’ 10
Running on Zero Agents Featured 453 DeepSeek OCR Demo π 453 An interactive demo for the DeepSeek-OCR model.
FreedomIntelligence/medical-o1-reasoning-SFT Viewer β’ Updated Apr 22, 2025 β’ 90.1k β’ 15.8k β’ 1.18k