Kishan Panaganti
kishanpb
AI & ML interests
LLM Reasoning via RL and anything RL
Recent Activity
upvoted a paper 7 days ago
FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience submitted a paper 7 days ago
FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience upvoted a paper 28 days ago
UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context DemonstrationsOrganizations
None yet