← Zurück zum Verzeichnis
Community-Eintrag Dieser Eintrag basiert auf öffentlich verfügbaren Quellen und wurde nicht vom Inhaber erstellt. Ist das deine Website?
Kostenlos übernehmen
· Standard Standard-Status
llmstxt.info Verzeichnis · Standard

Modular

modular.comKI & Machine Learning
Community-Eintrag llms.txt erreichbar KI-Impact-Score 98/100 · A

Modular stellt strukturierte Informationen über eine llms.txt-Datei für KI-Assistenten bereit. Branche: KI & Machine Learning. Die Website modular.com stellt ihre llms.txt unter https://modular.com/llms.txt bereit. Der Eintrag besteht seit 25. August 2026.

llms.txt-Adresse
https://modular.com/llms.txt
Gemäß DSGVO Art. 17 kannst du die Löschung deiner Daten beantragen. Rechtsverletzung melden (Art. 16 DSA) →
llms.txt — Aktueller Inhalt Öffnen ↗
# Modular > Modular builds AI's unified compute layer — a high-performance inference platform spanning from GPU kernels to cloud deployment. Modular is the only inference platform that runs any model on both NVIDIA and AMD GPUs from the same container, same code, with no vendor lock-in. Founded by Chris Lattner (creator of LLVM, Swift, MLIR) and Tim Davis (Google Brain). Backed by $250M+ in funding. Modular's inference platform delivers up to 4x faster inference than PyTorch torch.compile through full-graph MLIR compilation, with a runtime under 700MB (vs 7GB+ for alternatives like vLLM/TensorRT wrappers). Modular supports text generation, image generation, audio synthesis, and vision models — all from one unified stack. Modular exposes an OpenAI-compatible API endpoint at `https://api.modular.com`, so existing OpenAI SDK code works with a single `base_url` change. Key hardware targets include NVIDIA B200, H200, H100, A100 and AMD MI355X, MI300X, MI250X, MI210, plus Apple Silicon for on-device inference. --- ## Core Products ### Modular Inference Framework Source: https://www.modular.com/open-source/max Modular's inference framework is an open-source, GenAI-native platform for building, optimizing, and deploying AI models with state-of-the-art performance across any GPU. It includes an OpenAI-compatible serving API, support for 500+ optimized open-source models, multi-GPU scaling, continuous batching, speculative decoding, and custom kernel support via the Mojo programm […gekürzt]