Preference Model

Preference Model
Stealth ModePreference Model is building the next generation of training data to power the future of AI by developing machine learning infrastructure software for reinforcement learning experimentation.
About
What does Preference Model do?
Preference Model develops machine learning infrastructure software designed to support experimentation in reinforcement learning. They build high-quality RL training environments that reflect real-world complexity, with diverse tasks and robust reward functions, to teach models to perform ML research. The company partners with frontier AI labs to build capabilities for the next generation of LLMs.
Where is Preference Model headquartered?
Preference Model is headquartered in San Francisco, United States.
When was Preference Model founded?
Preference Model was founded in 2025.
What industry does Preference Model operate in?
Preference Model operates in Reinforcement Learning, Reinforcement Learning from Human Feedback, Synthetic Data, Synthetic Data Generation, MLOps, Developer Tools.
How many employees does Preference Model have?
Preference Model has approximately 17 people on record.
A Chinese AI company developing the Kimi chatbot and a series of large language models for various applications.
A benchmark for evaluating video world models as surrogate policy evaluators for embodied robot policies.
Frontier AI models compete in physics-driven combat and movement challenges.
Developing advanced AI models for complex problem-solving and scientific discovery.
Embodied builds AI agents that can interact with the real world through robotics and computer vision.
An entrant is a company tagged Reinforcement Learning whose domain was first registered in the window, counted from registry records in the StartupHub directory. 2 of them registered in the last 30 days. Registry detection runs two to three weeks behind registration, so recent weeks are a floor.
No comments yet. Be the first to share your take.


