PTENES
In-depth analysis · Agnes AI (Sapiens AI)

Text, images, and video for US$ 0 — is it worth it?

A multimodal API compatible with the OpenAI standard, with its main models available for free. Here’s what’s real, what the limits are, and what the risks are.

Agnes AI — text, images, video, and agents, 100% free
What it is

A free multimodal API — with fine print

Agnes AI, from Sapiens AI (Singapore), offers its own text, vision, image, and video models through an OpenAI-compatible API. “Free” means no charge per token or generation — not unrestricted or guaranteed use, or suitability for critical production.

💸 Exceptional cost

Text, images, and video currently at US$ 0 — including via API. One of the most aggressive free offers of 2026.

🔌 OpenAI-compatible

Same call format as OpenAI SDKs: integrate with n8n, Python, Node, agents, and existing apps by changing only the base URL and key.

⚠️ No SLA on the free plan

No uptime guarantee, limits may change, and on the free plan your data may be used to train models (unless you opt out). Don’t use it as your only provider in a critical system.

Models

The three main models

A single key and a single gateway (apihub.agnes-ai.com) for text/vision, images, and video.

💬 Agnes 2.0 Flash

Text, vision, and agents. Conversation, reasoning, code, tool calls, streaming, and vision via image URL. Declared context of 512K tokens (external catalogs cite 256K — test before sending very large contexts). US$ 0 per million tokens.

POST https://apihub.agnes-ai.com/v1/chat/completions
model: agnes-2.0-flash

🖼️ Agnes Image 2.1 Flash

Text→image, image→image, redesign, instruction-based editing, returns as URL or Base64, resolutions from 1K to 4K in 1:1, 16:9, 9:16, 4:3, 3:4, 2:3, 3:2, and 21:9 aspect ratios. US$ 0 per image.

POST https://apihub.agnes-ai.com/v1/images/generations
model: agnes-image-2.1-flash

🎬 Agnes Video V2.0

Text→video, image→video, keyframes, prompt-based camera/motion control, 480p/720p/1080p, asynchronous generation. Up to 441 frames per request (formula 8n+1). US$ 0 per second.

POST https://apihub.agnes-ai.com/v1/videos
GET  .../agnesapi?video_id=ID
Aspect ratio1K2K3K4K
1:11024×10242048×20483072×30724096×4096
16:91312×7362624×14723936×22085248×2944
9:16736×13121472×26242208×39362944×5248
4:31152×8642304×17283456×25924608×3456
3:4864×11521728×23042592×34563456×4608

Dimensions outside the table (e.g., 1920×1080) may be converted automatically. For a YouTube thumbnail: request "size":"2K","ratio":"16:9" and resize to 1920×1080.

How it works

Integration in minutes, layered architecture

Because the API is OpenAI-compatible, the usual approach works—and the safest architecture uses Agnes as a free, high-volume layer, never as the only brain.

Create an account + key→ Point the SDK to apihub.agnes-ai.com→ Call text / image / video→ Fallback to another provider

🅰️ Agnes layer

Free, high-volume tasks: batch thumbnails, summaries, classification, test agents, experimental videos.

🔁 Second option

Gemini / GLM / DeepSeek as an alternative route when Agnes is unavailable or limited.

🧠 Critical tasks

OpenAI / Claude for workloads that require predictable quality, an SLA, and broad validation.

Limits & plans

What the free plan offers—and what paid plans buy

The free plan is advertised as free indefinitely, but without a contractual guarantee: availability, quotas, and policies may change. Paid plans mainly buy speed and priority, not “unlimited use.”

🆓 Free plan

FeatureEffective limit
Text20 req/min
1K images20 img/min
2K images10 img/min
3K / 4K images1 img/min
Video1 req/min

Theoretically 28.800 text calls/day and 1.200 1K images/hour — but queues, fair use, and capacity reduce actual volume. No SLA.

💳 Token Plans (paid)

PlanPrice/month*Text / 5hText / week
StarterUS$ 41.50015.000
PlusUS$ 107.50075.000
ProUS$ 5030.000300.000

All Token Plans: up to 1.000 RPM for text, 4.000 images/day, 500 s of video/day, 100 RPM (1K) / 80 RPM (2K) for images, and 5 RPM for video. *Prices found in a community source — confirm at checkout. Odd detail: the paid plan increases speed but introduces explicit time-based quotas that aren’t published for the free plan.

You can combine: a free key and a Token Plan key use separate pools — when the paid quota runs out, the free key still works within the free limits. Creating multiple keys in the same category no multiplies limits (they share the same pool).

User guide · step by step

From key to first generation

Everything through the OpenAI-compatible endpoint. Change $AGNES_API_KEY with your key.

1

Create your account and API key

Sign up through the official campaign — agnes-ai.com/campaign — and generate a free key. Save it as an environment variable — never in the code.

# access: https://agnes-ai.com/campaign
export AGNES_API_KEY="sua-chave-aqui"
2

Text and agents — Agnes 2.0 Flash

OpenAI-format chat completions, with streaming and tool calls.

curl https://apihub.agnes-ai.com/v1/chat/completions \
  -H "Authorization: Bearer $AGNES_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"agnes-2.0-flash","messages":[{"role":"user","content":"Resuma o que é a Agnes AI em 3 frases."}]}'
3

Or use the OpenAI SDK pointed at the gateway

In Python (or Node), just change base_url — the rest of the code doesn’t change.

# pip install openai
from openai import OpenAI
client = OpenAI(base_url="https://apihub.agnes-ai.com/v1", api_key="$AGNES_API_KEY")
r = client.chat.completions.create(model="agnes-2.0-flash",
    messages=[{"role":"user","content":"Olá!"}])
print(r.choices[0].message.content)
4

Images — Agnes Image 2.1 Flash

E.g., 16:9 thumbnail in 2K (then resize to 1920×1080).

curl https://apihub.agnes-ai.com/v1/images/generations \
  -H "Authorization: Bearer $AGNES_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"agnes-image-2.1-flash","prompt":"thumbnail de YouTube sobre IA gratuita, estilo tech dark","size":"2K","ratio":"16:9"}'
5

Video — Agnes Video V2.0 (asynchronous)

Create the job, receive a video_id and poll until it’s ready. Frames follow the formula 8n+1 (max. 441).

# create
curl https://apihub.agnes-ai.com/v1/videos \
  -H "Authorization: Bearer $AGNES_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"agnes-video-v2.0","prompt":"drone sobrevoando cidade futurista ao entardecer","resolution":"720p"}'

# check the result
curl "https://apihub.agnes-ai.com/agnesapi?video_id=SEU_ID" \
  -H "Authorization: Bearer $AGNES_API_KEY"
6

Opt out of training (important!)

On the free plan, prompts and files may be used to train models unless you opt out—available within the service or through support. Never send sensitive data (CPF, patients, contracts, keys, LGPD data) through the free plan.

# support channel for opting out of training use
support@agnes-ai.com
Tips · measured in a real test

The 6 rules that matter financially

This isn’t theory: it came from ~70 real API calls, comparing what the documentation promises with what you have observed. The full record — accepted parameters, those that fail with HTTP 400, quota, cost, and open questions — is in NOTAS-API.md.

🇬🇧 Prompts in English

It’s not a matter of quality (they’re tied)—it’s the content filter: in Portuguese, it blocks legitimate generation with HTTP 400. Write the prompt in English, even for content in PT.

🦊 Ask a tail

Any animal with a tail in a front-facing pose comes out with two. The fix is to be explicit in the prompt: ONE SINGLE bushy tail. This applies to the general default — whatever you don’t specify, the model duplicates.

🔁 Retry with backoff is mandatory

About 34% of calls fail with 503 — and retries recover nearly 100%. Without retries, you may wrongly conclude that the API “doesn’t work.”

🎨 Use aesthetic terms only for style

Descriptors such as fur, expressive eyes or children's book inject characters in a prompt that was only meant to describe a setting. In the style description, use aesthetic terms only.

📐 1K for volume

1K takes ~32s; 4K takes ~153s and fails much more often. Resolution costs nothing extra (it’s all US$ 0), but costs time and increases the error rate — generate at 1K when volume matters.

⬇️ Download the PNG right away

The return URL is temporary. Save the file in the same step as generation; don’t keep the link assuming you can retrieve it later.

⚠️

Two pitfalls that falsely return HTTP 200

response_format at the JSON root level returns HTTP 400 — it only works inside extra_body. Worse: unknown parameters are silently discarded (the gateway uses drop_params), so a 200 OK no proves that your parameter was used. And there is no seed or fine-tuning/LoRA.

curl https://apihub.agnes-ai.com/v1/images/generations \
  -H "Authorization: Bearer $AGNES_API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"agnes-image-2.1-flash","prompt":"...","size":"1K","ratio":"16:9",
       "extra_body":{"response_format":"url"}}'
Performance

What the numbers say (and don’t say)

The evidence is strong for agents and reasonable for images; in video, the model is far from the leaders. Broad, independent text benchmarks (SWE-bench, GPQA, MMLU-Pro, etc.) are lacking.

🤖 Text/agents

Claw-Eval: Pass³ 60.9% — 9th place (snapshot leader: Claude Opus 4.6 at 70.4%). A legitimate benchmark of 300 real tasks verified by humans. Competitive, but not the leader — and a single benchmark does not prove overall superiority in coding, Portuguese, math, or factuality.

🖼️ Image

~1.178–1.184 Elo in editing evaluations: capable and competitive. Great for thumbnail alternatives, banners, backgrounds, and A/B testing; for premium assets (hands, typography, exact logos), review them or finish in a stronger model.

🎬 Video

Ranked around 61 (T2V) and 63 (I2V) — behind the leaders in realism and complex motion. Before it was free, it cost about US$ 0.30/min through gateways. At US$ 0, it’s worth testing extensively for supporting scenes and high volume.

Verdict

Is it worth it? Yes — with a fallback

The offer is too compelling to ignore for testing, content, agents, and non-critical automation. But don’t use it as the sole provider for a critical commercial system.

✅ Where to use it

  • Batch thumbnails, course and blog images;
  • Prototypes, agent testing, n8n automations;
  • Summaries, classification, free fallback for other APIs;
  • Experimental videos and supporting scenes;
  • Free INEMA.club applications and internal tasks without confidential data.

🚫 Where to avoid relying on it alone

  • Medical/dental, financial, and legal services;
  • Personal data protected (LGPD) on the free plan;
  • Applications with an SLA and agents that can’t go down;
  • Generation where the final quality needs to be predictable;
  • Commercial systems without a fallback.
Aplicação
   ↓
Roteador de modelos
   ├── Agnes AI: tarefas gratuitas e alto volume
   ├── Gemini/GLM/DeepSeek: segunda opção
   └── OpenAI/Claude: tarefas críticas ou complexas
STRONG
Cost, compatibility, and multimodalityUS$ 0 across three modalities, OpenAI standard, 512K context, 20 free RPM for text and 1K images, images up to 4K, good agent results (Claw-Eval).
WEAK
Validation, SLA, and continuityFew independent benchmarks, video quality trails the leaders, no SLA on the free plan, free-plan data may be used to train models (unless you opt out), promotional pricing without guarantees, a new company, and inconsistent docs (256K vs 512K; 30 vs 20 RPM).
RULE
High-capacity free tier — never your only brainUse Agnes for volume and experiments, with a model router and fallback to Gemini/GLM/DeepSeek and OpenAI/Claude for critical tasks.