Modular
We rebuilt the modern AI software stack, from the ground up, to boost any AI pipeline, on any hardware.
Modular is an AI platform designed to enhance any AI pipeline, offering an AI software stack for optimal efficiency on various hardware. It features popular models like Llama2, Mistral, StarCoder, and others, delivering performance and portability. Modular provides a suite of tools and libraries, including the MAX Engine for model inference and the Mojo programming language, enabling AI engineers to achieve throughput and cost savings while maintaining programmability and integration across different hardware environments.
Pricing: Usage-based
Resources
Modular Alternatives
Explore 40 products in the Frameworks & Stacks category. View all Modular alternatives.
llama.cpp
LLM inference in C/C++ with broad hardware support and aggressive quantization
vLLM
High-throughput LLM inference engine with PagedAttention for efficient GPU memory usage
Magnitude
Open-source desktop inference engine that tunes its kernels to your hardware and connects local models to your agent.
LiteLLM
Unified OpenAI-compatible proxy for 100+ LLM providers with cost tracking and load balancing
Work on Modular? Feature it at the top of Frameworks & Stacks.
Is your product missing?