Running 238 The ultimate guide to RL environments: building and scaling them in the LLM era π 238 Building and scaling RL environments for LLM training
Running on CPU Upgrade 280 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens π 280 Visualize syntheticβdata experiments as an interactive bookshelf
OpenMed/OpenMed-PII-BioClinicalModern-Large-395M-v1 Token Classification β’ 0.4B β’ Updated Jan 13 β’ 19.1k β’ β’ 10
Running on Zero Agents Featured 453 DeepSeek OCR Demo π 453 An interactive demo for the DeepSeek-OCR model.
FreedomIntelligence/medical-o1-reasoning-SFT Viewer β’ Updated Apr 22, 2025 β’ 90.1k β’ 21.1k β’ 1.19k