Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

A new model called Space Bunny (stealth model) is free on open router • 1M context • Text, image + video input • Adjustable reasoning • Tool calling OpenCode is also offering a free preview for this week with zero data retention. Try it on your next coding task....

16,948 Aufrufe • vor 3 Tagen •via X (Twitter)

27 Kommentare

Profilbild von Shruti
Shrutivor 3 Tagen

You can try it here:

Profilbild von Muhammad Ayan
Muhammad Ayanvor 3 Tagen

Tried it today, I got a 3D flight tracker from a sketch in 3.5 minutes :)

Profilbild von Khilesh | AI Tools
Khilesh | AI Toolsvor 3 Tagen

1M context is a huge advantage for serious coding tasks.

Profilbild von Mishika AI
Mishika AIvor 3 Tagen

Space Bunny being free on OpenRouter is already interesting.

Profilbild von Samuel Hu
Samuel Huvor 3 Tagen

1m context is wild, but the first thing i'd test is a 200k-token repo with one stale AGENTS.md and a screenshot assertion. fast demos hide where it starts dropping constraints. did you try a long tool-calling run or just the build prompt?

Profilbild von notgwapo
notgwapovor 3 Tagen

space bunny sounds wild, ngl. lowkey wanna try it out on my next project. 👀

Profilbild von Farhan Azad Shuvra
Farhan Azad Shuvravor 3 Tagen

That's impressive!

Profilbild von Ulubatlıhasanım.
Ulubatlıhasanım.vor 3 Tagen

Adjustable reasoning is probably the feature I’d experiment with first.

Profilbild von Ryan Carter
Ryan Cartervor 3 Tagen

1M context could be seriously useful when working across large codebases and multiple files.

Profilbild von Re seven
Re sevenvor 3 Tagen

Wonderful

Profilbild von Yasir Ai
Yasir Aivor 3 Tagen

1M context, multimodal input, adjustable reasoning, and tool calling in one free model is a pretty strong package for developers.

Profilbild von Justin Brave💡
Justin Brave💡vor 3 Tagen

Fantastic share

Profilbild von Daniel Jossr
Daniel Jossrvor 3 Tagen

Text, image, and video input in one model opens up some really interesting workflows.

Profilbild von AI VerseX🚀
AI VerseX🚀vor 3 Tagen

A stealth model suddenly becoming available for free is a pretty interesting move.

Profilbild von Mr. Jason💡
Mr. Jason💡vor 3 Tagen

This is the kind of AI thread that provides real practical value.

Profilbild von Holland Tech Marco
Holland Tech Marcovor 3 Tagen

Tools like this show how quickly AI is moving from experimentation into everyday workflows.

Profilbild von Dhruv kumar
Dhruv kumarvor 3 Tagen

Amazing share

Profilbild von Jara
Jaravor 3 Tagen

Wonderful

Profilbild von Mimu | AI Tools & News
Mimu | AI Tools & Newsvor 3 Tagen

Gave it a shot today turned a quick sketch into a full 3D flight tracker in literally 3.5 minutes.

Profilbild von Ashley Nicole
Ashley Nicolevor 3 Tagen

Awesome thanks

Profilbild von Shruti
Shrutivor 3 Tagen

Glad its helpful

Profilbild von Jeremy Longworth
Jeremy Longworthvor 3 Tagen

Stealth models are becoming a pretty good reminder not to overfit your workflow to a brand name. If a free model can take 1M context, reason, see images/video and call tools, the sensible move is to test it on your actual workload and keep whatever clears the bar.

Profilbild von AI Quanting
AI Quantingvor 3 Tagen

Any guesses on whose it is? Video input plus 1M context feels like a pretty short list.

Profilbild von John Smith
John Smithvor 3 Tagen

Done

Profilbild von Vipul Kumar Kewat
Vipul Kumar Kewatvor 3 Tagen

A free 1M-context model with multimodal input and tool calling is definitely worth testing on a real coding workflow.

Profilbild von frank z993
frank z993vor 3 Tagen

true no token limit on opencode

Profilbild von Evon
Evonvor 3 Tagen

Awesome post

Ähnliche Videos

HERMES AGENT SUPPORTS 300+ MODELS. PICKING THE RIGHT ONE PER TASK IS THE DIFFERENCE BETWEEN $5/MONTH AND $50. STARTING OUT: Claude Sonnet 4.6. official recommendation from Nous Research. "the model this project was built and tested with." strong reasoning. reliable tool calling. mid-range pricing. PREMIUM TIER: Claude Opus 4.8. best coding benchmarks available. self-correcting reasoning. catches its own mistakes. 1M context. use for demanding tasks where quality matters. GPT-5.5. #1 Chatbot Arena. #1 GPQA Diamond reasoning (94.1%). #1 creative writing. 2M context. handles entire codebases in one pass. Grok 4.30. the only frontier model with live X firehose access. real-time social data, breaking news, market sentiment. connects via Grok OAuth. no separate API key. Grok-Composer-2.5-Fast (v0.17.0). Cursor's coding model. 200K context. available through your Grok subscription via OAuth. no extra cost if you already pay for Grok. MID-RANGE TIER: Claude Sonnet 4.6. best balance of quality and cost for daily use. strongest prose and tool calling in this tier. Gemini 2.5 Pro. Google Search grounding built in. cites sources. verifies claims. pulls current data. 2M context. best for research-heavy workflows. GPT-4.1. reliable tool calling. solid general reasoning. good middle ground when you need OpenAI compatibility. BUDGET TIER: Claude Haiku 4.5. fastest Anthropic model. cheapest paid Claude option. strong at classification, routing, simple queries. use for auxiliary tasks: compression, vision, web extraction, approval scoring. DeepSeek V4. best cost-to-quality ratio in the market. 90% cache discount on repeated context. use for sub-agents and bulk parallel work. DeepSeek V4 Flash. cheapest paid model worth using. 1M context. MIT license. self-hostable. use for cron jobs, monitoring, routine searches. MiniMax M3. Nous Research and MiniMax collaborating on optimization. 1M context via lightning attention. 59% SWE-Bench Pro. beats several premium models on coding. one of the most-used models inside Hermes. FREE / LOCAL: Qwen 3.5 27B via Ollama. 16GB VRAM. reliable tool calling. best free local model for Hermes as of mid-2026. Qwen 3 8B. 8GB VRAM. fits a $7 VPS. handles routine tasks at zero API cost. Llama 4 Maverick. best open-weight tool calling. 1M context. needs more VRAM but strongest local option. HOW TO ASSIGN MODELS: main model: Desktop app / Dashboard → Models → switch sub-agent model: set in Desktop app, Dashboard, or config.yaml: delegation: model: "deepseek/deepseek-v4" auxiliary models (compression, vision, web extract): Desktop app / Dashboard → Models → Auxiliary Haiku 4.5 or Gemini Flash work well here. saves significantly when your main model is premium. per-profile: each Hermes profile gets its own model. Scout on DeepSeek. Analyst on Sonnet. Briefer on budget model. Coder on Opus. per-cron-job: pin a specific model to any cron job. morning brief on Haiku. deep research on Sonnet. monitoring on DeepSeek Flash. each job uses only the model it needs. per-session: /model deepseek/deepseek-v4-flash hot-swap mid-conversation. no restart needed. FALLBACK CHAINS: if your primary model is unavailable, Hermes automatically switches to the next provider. rate limit or server error = next model in the chain. no failed runs. no manual intervention. set in Desktop app, Dashboard, or config.yaml: fallback_providers: - openrouter - nous - codex PROVIDER PATHS: OPENROUTER: 300+ models under one API key. pay per token. most flexible. NOUS PORTAL: 300+ models + Tool Gateway (web search, image gen, TTS, browser). one OAuth. one subscription. 10% off token-billed providers. CHATGPT SUB: GPT-5.5 + Grok via OAuth. included tokens with $20 subscription. OLLAMA: free. local. private. zero API cost. your hardware only. mix providers across profiles and tasks. Scout on OpenRouter. Analyst on Nous Portal. Coder on ChatGPT sub. Monitor on Ollama. THE RULE: premium for work that needs deep reasoning. mid-range for daily driver tasks. budget for volume and background work. free for monitoring and routine jobs. pricing changes fast. check openrouter ai for current rates before committing. Which is your favourite model and for what task? full 15 levels breakdown in the article 👇

YanXbt

17,138 Aufrufe • vor 3 Monaten