solved classic rl environments
💼 Hiring
Nitish Pandey
nitishpandey04
AI & ML interests
LLMs, Translation
Recent Activity
commentedon a paper about 10 hours ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement upvoted a paper about 10 hours ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement commentedon a paper 4 days ago
Kimi K3: Open Frontier Intelligence