Managed compute
Choose CPU or GPU resources for an endpoint. Review the price before starting and configure an idle timeout.
Modal compute endpoints
Explore endpointsEdelux Foundry
Models, agents and applications in one workspace, with your infrastructure and your budgets under control.
Prompt + knowledge
Agent + tools
Connect a model to your knowledge and tools. Inspect each run before putting it into a workflow.
Build an agentModel registry
Open-weight releases, connected to the serving engine your infrastructure can actually run.
55 matching models
2.8T MoE, 104B active, 1M context
Kimi K3 License
2.4T MoE, 95B active, 262K context
Qwen3.8-Max License
1.6T MoE, 49B active, 1M context
MIT
753B MoE, 40B active, 1M context
MIT
428B MoE, 23B active, 1M context
MiniMax Community License
284B MoE, 13B active, 1M context
MIT
Model links open upstream weights. Inclusion in this roster does not mean a hosted endpoint is running.
Use managed compute for experiments, or take a deployment bundle to your own infrastructure. Keep the deployment artifact either way.

Choose CPU or GPU resources for an endpoint. Review the price before starting and configure an idle timeout.
Modal compute endpoints
Explore endpointsDownload a portable bundle, review its configuration, and apply it to a Kubernetes cluster, Compose host or Linux server.
Kubernetes · Docker · Linux
Explore bundlesThe setup script detects the coding tools already installed on your machine and writes the gateway endpoint, token and model profiles into each one.
Control who can run workloads, what they can spend, and which steps need a human decision.

Sign in to deploy a model, connect a cluster, and see real usage and cost per team.