Geodd
Managed AI inference endpoints and GPU infrastructure with OpenAI-compatible API
Geodd is an AI inference platform offering serverless endpoints, dedicated inference, and GPU clusters for production workloads. The API is OpenAI SDK-compatible, so switching providers requires changing one line of code. Geodd applies inference optimizations at the model and runtime layers (custom CUDA kernels, disaggregated prefill/decode, KV cache routing, FP8/FP4 quantization) to increase throughput without hardware upgrades. The platform claims 25-50% more throughput on existing GPU fleets and 2-3x faster generation via adaptive speculative decoding. Primary region is North America East (500+ GPUs), with EU and APAC regions coming.
Pricing: Per token usage
Geodd Alternatives
Explore 88 products in the Inference APIs category. View all Geodd alternatives.
Infer by Flow7
Responses API gateway for coding agents, with a public model catalog and per-request spend ceilings
TensorX
EU-sovereign inference API with 42+ open-source models and zero data retention
EUrouter
European AI gateway that routes to 100+ models with EU data residency
Lyceum
EU-hosted inference cloud for open-source models, OpenAI-compatible
Project Zero
CPU-only LLM inference engine in C with no runtime dependencies
Akumi
EU-hosted OpenAI-compatible inference API with data residency, PII pseudonymization and per-request audit trails
Work on Geodd? Feature it at the top of Inference APIs.
Is your product missing?