🔌 LLM APIs with free tier · no card · provisional — added recently, two weeks of probes still to pass · live — last verified by a probe on 2026-09-17 · yolo-auto.com · back to the whole list
One model, Qwen3.8 Flash — Qwen’s open-weight Qwen3.8-Flash-Next — served on the vendor’s own flat-rate API for coding agents; the free plan is a small daily allowance with no card
qwen3.8-flash
“15 free requests a day. No card required.” on the home page and “No card required, free forever” on the Free plan card, at 128K context, against $19/mo Builder and $39/mo Pro — a handful of agent turns a day. The model is Qwen3.8 Flash, id qwen3.8-flash, charted with Artificial Analysis scores for Qwen3.8-Flash-Next, the mixture-of-experts Qwen published open-weight on 2026-08-24 (about 180B parameters, 10 of 512 experts active). Yolo-Auto “runs the model-serving stack rather than reselling a third-party model API”. The FAQ calls the free tier “for testing” where the plan card says “free forever”, and the terms forbid using “multiple accounts … to combine capacity”. Sign-in is through Google, GitHub or Discord, and prompt and response bodies are “not routinely retained”. Read 2026-09-17
https://yolo-auto.com/v1YOLO_AUTO_API_KEY — get one at https://yolo-auto.com/appqwen3.8-flash15 free requests a day, No card required, free forever2026-09-17 — Free models changed: added qwen3.8-flash; dropped qwen3.82026-09-14 — Added to the list: One model, Qwen3.8-27B in FP8, served on the vendor’s own flat-rate API for coding agents; the free plan is a small daily allowance with no cardGenerated from registry.yaml on 2026-09-18 and re-verified twice a week; the full list, the Atom feed and the machinery are at https://github.com/mvalentsev/awesome-free-ai-coding.