Neuralmagic

NeuralmagicNeuralmagic

Neuralmagic

Stealth modeStealth Mode

Software-delivered AI optimization platform for high-performance, cost-efficient deep learning inference on commodity CPUs.

DR 0Speed 692026Active
Rate

About

Neural Magic provides software-based model optimization and compression technologies that enable enterprises to run deep learning models, including large language models (LLMs), on commodity CPUs with GPU-class performance. The company is a leading contributor to the open-source vLLM project and focuses on reducing the cost and complexity of AI deployment across hybrid cloud environments.
Frequently asked

What does Neuralmagic do?

Neural Magic provides software-based model optimization and compression technologies that enable enterprises to run deep learning models, including large language models (LLMs), on commodity CPUs with GPU-class performance. The company is a leading contributor to the open-source vLLM project and focuses on reducing the cost and complexity of AI deployment across hybrid cloud environments.

Is Neuralmagic trustworthy and reputable?

StartupHub's Data Trust & Reputation score for Neuralmagic is 73 out of 100, based on site security posture and privacy practices.

When was Neuralmagic founded?

Neuralmagic was founded in 2026.

What industry does Neuralmagic operate in?

Neuralmagic operates in AI Infrastructure, MLOps, Generative AI, Large Language Model, Edge AI, Enterprise Software.

New entrants in MLOps
Last 90 days
Comments
(6)
6 positive0 mixed0 negative
Reddit
r/machinelearningnewsu/ai-loverNov 25, 2024Positive

Neural Magic has responded to these challenges by releasing Sparse Llama 3.1 8B—a 50% pruned, 2:4 GPU-compatible sparse model that delivers efficient inference performance.

View on Reddit
Reddit
r/machinelearningnewsu/ai-loverAug 31, 2024Positive

GuideLLM Released by Neural Magic: A Powerful Tool for Evaluating and Optimizing the Deployment of Large Language Models (LLMs).

View on Reddit
Reddit
r/InfermaticAIu/InfermaticAug 5, 2024Positive

WELCOME TO: neuralmagic/Meta-Llama-3.1-405B-Instruct. With 16K context, this chunky model is available now on the UI and also on the API.

View on Reddit
Reddit
r/InfermaticAIu/InfermaticAug 2, 2024Positive

Our dynamic models have received a performance boost! The speed and stability of Magnum, Euryale, Miquliz, Tenyx, and more are now unmatched. Enjoy faster and more reliable performance.

View on Reddit
Reddit
r/MachineLearningu/markurtzMay 21, 2024Positive💎

In a collaboration across Neural Magic, Cerebras, and IST Austria, we've pushed out, to the best of our knowledge, the first highly sparse, foundational LLMs with full recovery on several fine-tuning tasks, including chat, code generation,…

View on Reddit
Reddit
r/machinelearningnewsu/ai-loverMay 18, 2024Positive

Researchers from Cerebras & Neural Magic Introduce Sparse Llama: The First Production LLM based on Llama at 70% Sparsity.

View on Reddit

Some comments are pulled from public discussions around the web (look for the source icon). Quotes are excerpts; click through to read the full thread.