profundo.ai

PRICING

Frontier open-weight models. Priced for monthly use.

Profundo gives you direct access to a growing catalog of open-weight frontier models — GLM 5.2, MiniMax M3, and Nemotron Ultra. Pick the monthly plan that matches your workload. No per-token meter. No surprise overage.

FOUNDER

$20

/month

Launch-period members building on Profundo.

  • 30,000 requests/month (Priority-level access)
  • All available models included
  • Chat Completions + Responses API
  • Native tool calling + parallel tool calls
  • Priority roadmap input
  • Locked-in rate for founding cohort

30,000 requests/month, may increase as demand grows

Become a Founder

PLUS

$5

first month, then $10/mo

For everyday frontier-model use.

  • All available models, 250 requests/day
  • Chat Completions + Responses API
  • Native tool calling + parallel tool calls
  • Same frontier models as Priority
  • Generous cap covers most workflows

250 requests / day, hard cap

Get Plus

PRIORITY

$20

first month, then $40/mo

Heaviest workloads, agents, and automation.

  • 30,000 requests/month, fair-use
  • Chat Completions + Responses API
  • Native tool calling + parallel tool calls
  • Peak-ramp tolerance for agent loops
  • Best for Hermes / 24×7 workloads

30,000 requests/month — we may raise this as demand grows

Go Priority

Plus vs Priority

Standard tier plans, side by side. Founder members get Priority-level access at the founding rate.

PlusPriority
Monthly price$10/mo ($5 first month)$40/mo ($20 first month)
API accessChat Completions + ResponsesChat Completions + Responses
Tool callingYes, native + parallelYes, native + parallel
Monthly request cap250 requests / day (~7,500/month)30,000 requests / month
Same frontier modelsGLM 5.2, MiniMax M3, Nemotron UltraGLM 5.2, MiniMax M3, Nemotron Ultra
Agent loop toleranceStandardEnhanced peak-ramp tolerance
Best forEveryday use, standard workflows24×7 agents, heaviest workloads

What is included

Access covers the documented text Chat Completions and Responses API routes for all available models (glm-5.2, minimax-m3, nemotron-ultra), including native tool calling, parallel tool calls, structured output (response_format), and configurable reasoning effort. There is no public SLA and no guarantee of throughput, uptime, concurrency, or model availability.

Founders keep the founding rate while continuously subscribed. Plus is a hard cap of 250 requests/day. Priority includes 30,000 requests per month — we may raise this cap as demand grows.

Create your account