Dipankar Sarkar PRO
dipankarsarkar
AI & ML interests
Building the AI-native stack. Agents as infrastructure, safety as architecture, performance as plumbing. I publish the receipts: papers, datasets, demos.
Recent Activity
reacted to sergiopaniego's post with ๐ค 42 minutes ago
quick reminder! ๐จ
tomorrow (Tuesday, July 28), we're back with Class 3 of the Training Agents live series
๐ง what: reinforcement learning for training agents (GRPO): how it works, how to implement it in TRL, and end-to-end examples
๐๏ธ when: Tuesday, July 28 - ๐ 5:00 PM CEST / 8:30 PM IST
๐ where: Live on @huggingface's X, YouTube, and LinkedIn
live: https://www.youtube.com/watch?v=ztdTed5egrM
class 1: https://x.com/SergioPaniego/status/2069382207618379813
class 2: https://x.com/SergioPaniego/status/2075180665184686187 upvoted a paper about 2 hours ago
Progress Reward Modeling for Robotic Learning: A Comprehensive Survey