[metadata]
author: FlexAI
description: Apply to the FlexAI Startup Program for $1,000 in Token Factory credit, discounts, and $10/month in free credits for your first 3 months. Grow into dedicated GPU compute on one account as your workload scales.
og:description: Apply to the FlexAI Startup Program for $1,000 in Token Factory credit, discounts, and $10/month in free credits for your first 3 months. Grow into dedicated GPU compute on one account as your workload scales.
og:image: https://flex.ai/og-image.png
og:image:alt: FlexAI: The platform for agent-native AI
og:image:height: 630
og:image:width: 1200
og:title: Start building on FlexAI | Startups
og:type: website
og:url: https://flex.ai/startups
twitter:card: summary_large_image
twitter:description: Apply to the FlexAI Startup Program for $1,000 in Token Factory credit, discounts, and $10/month in free credits for your first 3 months. Grow into dedicated GPU compute on one account as your workload scales.
twitter:image: https://flex.ai/og-image.png
twitter:site: @FlexAI
twitter:title: Start building on FlexAI | Startups
viewport: width=device-width, initial-scale=1.0

[canonical-links]
https://flex.ai/startups

[document-links]
/
1 Model calls Token Factory 20+ models · one API key: /token-factory
2 Agent loops Agent SDK Tools · approvals · audit: /agent-sdk
3 Dedicated scale Dedicated Endpoints Reserved GPUs · same API: /dedicated-endpoints
4 Private AI cloud AI Factory VPC · on-prem · air-gapped: /ai-factory
AI Factory: /ai-factory
Acceptable Use Policy: /acceptable-use-policy
Agent SDK: /agent-sdk
All Models: /models
All use cases: /use-cases
Approach: /why-us
Blog: /blog
Blueprints: /blueprints
Builders & Startups: /startups
Careers: /careers
Case Studies: /case-studies
Dedicated Endpoints: /dedicated-endpoints
DeepSeek: /models/deepseek
Docs: /docs
Docs: https://docs.flex.ai
Embeddings: /embeddings
Engineering: /engineering
Enterprise: /enterprise
Explore Token Factory: /pricing
Fine-Tuning: /fine-tuning
GLM: /models/glm
GPT-OSS: /models/gpt-oss
Gemma: /models/gemma
Get an API key: https://platform.flex.ai/signup
Growing AI Teams: /companies
Image Generation: /image-generation
Inference: /inference
Llama 3.1 8B Instruct: /models/llama-3-1-8b-instruct
Llama: /models/llama
Migration Guide: /migrate
Mistral: /models/mistral
Nemotron: /models/nemotron
Partnerships: /partnerships
Platform Overview: /platform
Platform: /platform
Pricing: /pricing
Privacy Policy: /privacy-policy
Qwen3 30B A3B Thinking 2507: /models/qwen3-30b-a3b-thinking-2507
Qwen3 Coder 30B A3B: /models/qwen3-coder-30b-a3b
Qwen: /models/qwen
See live per-model performance.: /models
See pricing.: /pricing
See the pipeline: /use-cases/coding-agents
See the pipeline: /use-cases/research-agents
See the pipeline: /use-cases/support-agents
Sign Up: https://platform.flex.ai/signup
Speech-to-Text: /speech-to-text
System Integrators: /system-integrators
Talk to us: /contact
Terms of Service: /terms-of-service
Text-to-Speech: /text-to-speech
Token Factory: /token-factory
Tools & Calculators: /tools
Training: /training
Trust Center: https://security.flex.ai
Use Cases: /use-cases
Video Generation: /video-generation
https://security.flex.ai

[structured-data]
{"@context":"https://schema.org","@type":"FAQPage","mainEntity":[{"@type":"Question","acceptedAnswer":{"@type":"Answer","text":"$10/month in free credits for your first 3 months. Accepted Startup Program startups get $1,000 in Token Factory credit on top."},"name":"What free credits do I get?"},{"@type":"Question","acceptedAnswer":{"@type":"Answer","text":"Yes. A card is required to activate your account. You won't be charged while your free credits last."},"name":"Is a credit card required?"},{"@type":"Question","acceptedAnswer":{"@type":"Answer","text":"AI-native startups with a continuous inference or agent workload in production: typically $1M+ venture- or accelerator-backed, with at least one ML/AI engineer on the team."},"name":"Who's eligible for the FlexAI Startup Program?"},{"@type":"Question","acceptedAnswer":{"@type":"Answer","text":"$1,000 in Token Factory credit, then 50% off your first 3 months and 30% off the next 3. As your workload grows, you qualify for 50% off dedicated GPU and fine-tuning prices for 6 months, with committed-use pricing after that."},"name":"What do accepted startups get?"},{"@type":"Question","acceptedAnswer":{"@type":"Answer","text":"Book an intro call below. We'll walk through your workload, confirm fit, and get you onboarded, usually within 10 business days."},"name":"How do I apply?"},{"@type":"Question","acceptedAnswer":{"@type":"Answer","text":"No. FlexAI is OpenAI-compatible everywhere. You can leave with a config change."},"name":"Is there a contract or lock-in?"},{"@type":"Question","acceptedAnswer":{"@type":"Answer","text":"Competitive rates. The discount ending isn't a cliff, because your published rate is already competitive on every model."},"name":"What does it cost after the program?"}]}
{"@context":"https://schema.org","@type":"Service","areaServed":"Worldwide","description":"OpenAI-compatible inference with Token Factory credit and discounts for AI-native startups, scaling into dedicated GPUs on one account.","name":"FlexAI Startup Program","provider":{"@type":"Organization","logo":"https://flex.ai/flexai-logo.png","name":"FlexAI","url":"https://flex.ai"},"serviceType":"Managed AI platform","url":"https://flex.ai/startups"}

[content]
Start building on FlexAI | Startups
Skip to content
Products
Products
Token Factory
Dedicated Endpoints
Agent SDK
AI Factory
Platform
Platform Overview
Inference
Fine-Tuning
Training
Who We Serve
Builders & Startups
Growing AI Teams
Enterprise
System Integrators
Models
Browse
All Models
Use Cases
By family
Qwen
Mistral
DeepSeek
Llama
GLM
Nemotron
Gemma
GPT-OSS
By capability
Embeddings
Speech-to-Text
Text-to-Speech
Image Generation
Video Generation
Pricing
Resources
Docs
Blog
Blueprints
Case Studies
Migration Guide
Tools & Calculators
Company
Approach
Engineering
Partnerships
Careers
Trust Center
Talk to us
Sign Up
For startups
Start today.
Scale on the same account
Start with OpenAI-compatible inference, keep your agent stack portable, and move into dedicated GPUs when traffic proves out.
Built for AI-native startups shipping agents and multimodal products.
Get an API key
Apply to the Startup Program
Start building in minutes
Point the OpenAI SDK at FlexAI, swap the model, and ship. Free credit covers your first calls.
curl
Python
TypeScript
Model
Mistral Nemo
Gemma 4 26B A4B
Qwen3.6-27B
Nemotron 3 Super 120B A12B
GPT-OSS 20B
Qwen3.6-35B-A3B
Llama 3.1 8B Instruct
Qwen3.5 9B
Gemma 4 31B IT
GPT-OSS 120B
Qwen3 Coder 30B A3B
Llama 3.3 70B Instruct
Qwen3 8B
Qwen3 30B A3B Thinking 2507
MiniMax M2.7
GLM 5.2
Laguna S 2.1
GLM 4.5 Air
DeepSeek V4 Flash 0731
Muse Glimmer 30B
Copy
curl https://api.flex.ai/v1/chat/completions \ -H "Authorization: Bearer $FLEXAI_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "Meta-Llama-3.1-8B-Instruct-FP8", "messages": [{"role": "user", "content": "Hello from FlexAI"}] }'
50%
Off dedicated and fine-tuning prices for 6 months
$1,000
Token Factory credit in the program
20+
Open models behind one API key
Built for what you're shipping
The agents startups ship most, each a multi-model pipeline behind one key.
Coding agents
Agents that generate, review, and repair code across your repo.
Generate
Qwen3 Coder 30B A3B
Review & fast edits
Qwen3.6-35B-A3B
Start with
Qwen3 Coder 30B A3B
from
$0.07
/M
See the pipeline
Support agents
Agents that triage, answer, and resolve customer conversations.
Transcribe
Whisper Large V3 Turbo
Read attachments
PaddleOCR-VL 1.5
Retrieve
BGE-M3
Respond
Llama 3.1 8B Instruct
Speak
Kokoro-82M
Start with
Llama 3.1 8B Instruct
from
$0.02
/M
See the pipeline
Research agents
Agents that retrieve, reason over, and synthesize large source sets.
Retrieve
BGE-M3
Reason
Qwen3 30B A3B Thinking 2507
Summarize
DeepSeek-V4-Flash
Start with
Qwen3 30B A3B Thinking 2507
from
$0.08
/M
See the pipeline
All use cases
Why startups start on FlexAI
Ship on day one, on a platform that stays cheap as you scale.
Serverless, per-token
Start instantly: pay only for the tokens you use, with no provisioning and no minimums.
Performance
Streaming on every model, with low-latency serving tuned per model.
See live per-model performance.
Competitive cost
Competitive per-token pricing on every model, tracked to the market rate.
See pricing.
Already shipping agents? Bring them to the
Agent SDK
(in trial): portable skills and multi-model routing on the same key.
The FlexAI Startup Program
Apply-only, one program, three stages. $1,000 in Token Factory credit to start, 50% off your first 3 months and 30% off the next 3, then up to $20K of dedicated savings as you grow. You're never asked to re-apply or migrate.
Stage 1
Token Factory entry
$1,000 in Token Factory credit: our serverless per-token tier
Then 50% off your first 3 months, 30% off the next 3
Competitive pricing for the life of your account
Stage 2
Dedicated compute
As your workload proves out, you qualify for dedicated GPU-hours
50% off published dedicated and managed fine-tuning prices for 6 months, up to $20K of savings
One account and API as you scale
Stage 3
Committed use
Lock in committed-use pricing when you're ready
A designed continuation, not a second cliff
No minimum spend
Apply-only. We look for AI-native teams with a real, continuous workload:
$1M+ venture- or accelerator-backed, or a validated production workload
AI-native product: inference is central, not a side feature
At least one ML/AI engineer on staff
A continuous inference or agent workload in production, not project-based
Loading meeting scheduler…
Not there yet?
Explore Token Factory
and apply when your workload is ready. Reviewed within 10 business days.
One account,
the whole way
No re-platform when your workload grows.
1
Model calls
Token Factory
20+ models · one API key
2
Agent loops
Agent SDK
Tools · approvals · audit
3
Dedicated scale
Dedicated Endpoints
Reserved GPUs · same API
4
Private AI cloud
AI Factory
VPC · on-prem · air-gapped
Frequently Asked Questions
What free credits do I get?
$10/month in free credits for your first 3 months. Accepted Startup Program startups get $1,000 in Token Factory credit on top.
Is a credit card required?
Yes. A card is required to activate your account. You won't be charged while your free credits last.
Who's eligible for the FlexAI Startup Program?
AI-native startups with a continuous inference or agent workload in production: typically $1M+ venture- or accelerator-backed, with at least one ML/AI engineer on the team.
What do accepted startups get?
$1,000 in Token Factory credit, then 50% off your first 3 months and 30% off the next 3. As your workload grows, you qualify for 50% off dedicated GPU and fine-tuning prices for 6 months, with committed-use pricing after that.
How do I apply?
Book an intro call below. We'll walk through your workload, confirm fit, and get you onboarded, usually within 10 business days.
Is there a contract or lock-in?
No. FlexAI is OpenAI-compatible everywhere. You can leave with a config change.
What does it cost after the program?
Competitive rates. The discount ending isn't a cliff, because your published rate is already competitive on every model.
Ready to get started?
Book an intro call to join the Startup Program: $1,000 in Token Factory credit, stacked discounts, and free monthly credits.
Apply to the program
AI infrastructure that adapts as you grow.
Products
Token Factory
Dedicated Endpoints
Agent SDK
AI Factory
Platform
Pricing
Who We Serve
Builders & Startups
Growing AI Teams
Enterprise
System Integrators
Models
All Models
Use Cases
Resources
Docs
Blog
Blueprints
Case Studies
Migration Guide
Tools & Calculators
Company
Approach
Engineering
Partnerships
Careers
Trust Center
©
2026
FlexAI. All rights reserved.
Terms of Service
Privacy Policy
Acceptable Use Policy
