Kishan Panaganti
kishanpb
AI & ML interests
LLM Reasoning via RL and anything RL
Recent Activity
upvoted a paper 3 days ago
UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations upvoted a paper about 1 month ago
Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning authored a paper about 1 month ago
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and EditableOrganizations
None yet