NVIDIA B200
Latest-gen accelerators, dedicated, on bare metal. No shared tenant.
Token Factory for the API, Sandbox to build in, and dedicated GPUs for the heavy lifting.One balance, in euros, in a European data centre.
Activate with your card and you're up in under a minute. Use it for three minutes, pay for three minutes.
products
Inference with the API, agents in the Sandbox, heavy workloads on a dedicated GPU, separately or combined. All on a single balance, one invoice.
An AI that answers, via API
Connect your website, shop or application to the best open models. Compatible with the OpenAI and Anthropic SDKs.
Real example
a clinic answering appointment questions at any hour.
Your workspace in the cloud
A powerful machine switched on for you, with agent templates ready. It suspends itself when you stop using it, and your files stay put.
Real example
a team that builds and tests without depending on its own hardware.
A whole machine for you
A full NVIDIA B200 or a MIG slice, shared with nobody. By the month or by the year, for serious workloads.
Real example
a company training its own model on private data.
The path
One balance covers all three stages, with no need to switch providers.
Test the idea
Token Factory
Call a model from your app through the API. The free €50 goes a long way.
Build it
Sandbox
A persistent environment to build in, with agent templates ready to go.
Take it to production
Dedicated GPU
Dedicated power for your real workload, by the month or by the year.
Two questions
Two questions. Whether you run infrastructure for a living or have never touched AI before, we will tell you where to start.
The newest models on the market, already running on our cluster.
infrastructure
Six reasons to bring your AI here: what hardware runs it, where it sits, how it connects, where your data lives, which certifications it holds and how it is invoiced.
Latest-gen accelerators, dedicated, on bare metal. No shared tenant.
Data center on Spanish soil. No international transfers, no sub-processors outside the EU.
Low latency between nodes for multi-GPU and sustained throughput in production.
High-speed parallel storage for datasets, checkpoints and persistent volumes.
Operational certifications. Standard DPA ready to sign. Optional zero retention for enterprise.
Deductible for EU companies and freelancers, valid for intra-EU reverse charge. No USD/EUR conversion, no month-end FX surprises.
models
Open models from several providers, all deployed on our own infrastructure in Spain.
Text and code
Chat, writing and code, with tool calling.
Reasoning
Models that think before they answer, with adjustable effort.
Vision and documents
Reading images, PDFs and screenshots, with OCR and document understanding.
Voice
Speech to text and text to speech, including voice cloning.
Music
Full songs from lyrics and a style prompt.
Support
You send an email and a person replies, within four working hours. For critical incidents, a call with an expert.
Write to usfrequently asked questions
All data is processed at a datacenter in Spain. It never leaves the European Union and is not replicated to other regions. The platform is ISO 27001 and ENS Medium certified, and we have a standard DPA ready to sign.
Yes. The API exposes /v1/chat/completions (OpenAI) and /v1/messages (Anthropic). Point your SDK at api.gpuflow.ai with your API key and the catalog models are available directly. We support SSE streaming and function calling where the model allows it.
Account yes, 1 minute with email or SSO. Add your card and get 50 € of credit to call the API or spin up a Sandbox. The 1 € verification charge is refunded.
Prepaid model. Top up credit when you want and consume by real usage, measured in tokens (Token Factory) or active hours (Sandboxes). Automatic notification when balance drops below 20%. Deductible Spanish invoice on every top-up.
We only charge for the hours the sandbox is active. After the idle time you configured (from 10 min to 4 h, or manual stop only) with no SSH activity or user processes, it suspends automatically and stops billing compute. Your persistent volume survives intact and is available again in 5-8 seconds on reconnect.
Yes, anytime from the dashboard. Billing is prorated to the exact moment of cancellation. If you delete, the persistent volume and all snapshots are removed; you have up to 7 days to undo the deletion if it was accidental.
Yes, we publish real-time status at gpuflow.ai/status. The contractual SLA with support response times is agreed in enterprise plans.
That the service is already in production and billed on real usage, but we are still polishing the platform: there may be the occasional interruption and the interface changes from one week to the next. Prices, your balance and data residency in Spain do not change because of the beta. The state of every service is published on the status page, and if something breaks you write to us at [email protected] and we sort it out.
Data that never leaves Europe, billing in euros and one balance across API, Sandbox and GPU. Create an account, add your card and start with €50 free.
Building a European startup? Apply for up to €5,000 in credits