Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

Patronus AI

Team
company
Verified
https://patronus.ai
patronusai
Activity Feed Request to join this org

AI & ML interests

LLM Evaluation

Recent Activity

akkikiki  updated a model 2 days ago
PatronusAI/kimi-k3-nvfp4
anandnk24  published a model 3 days ago
PatronusAI/kimi-k3-nvfp4
DarshanDeshpande  new activity 12 days ago
PatronusAI/world_model_corpus:Add reinforcement-learning task category, paper and GitHub links
View all activity

Papers

Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Agentic RL

Benchmarking Reward Hack Detection in Code Environments via Contrastive Analysis

View all Papers

Rebecca Qian's profile picture Anand Kannappan's profile picture Bartosz Mielczarek's profile picture Bartosz Mielczarek's profile picture Varun Joshi's profile picture Arek's profile picture Darshan Deshpande's profile picture Maciej Gełdon's profile picture Shivani Jain's profile picture Varun Gangal's profile picture Edgar Colque's profile picture Jedrzej's profile picture Chinmayee Kulkarni's profile picture Devanshu Bansal's profile picture Bartlomiej Olechno's profile picture Tobi Akomolede's profile picture Yoshinari Fujinuma's profile picture Christopher Babayan's profile picture Nick Saban's profile picture Charvi Bannur's profile picture Mariya Vasileva's profile picture Zhe Li's profile picture
PatronusAI 's papers 3
Submitted by
Darshan Deshpande
7

Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Agentic RL

PatronusAI Patronus AI
2 2
Submitted by
Darshan Deshpande
1

Benchmarking Reward Hack Detection in Code Environments via Contrastive Analysis

PatronusAI Patronus AI
3
Submitted by
Darshan Deshpande
4

MEMTRACK: Evaluating Long-Term Memory and State Tracking in Multi-Platform Dynamic Agent Environments

PatronusAI Patronus AI
2
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs