Loading video...

Video Failed to Load

Go Home

Your OpenCode will no longer hallucinate. It now automatically detects when it needs docs, repos, or research papers, then indexes and fetches them via Nia. All retrieved context remains stateful. Introducing the Nozomio Labs opencode plugin. bunx nia-opencode@latest install

116,303 views • 6 months ago •via X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos

QVAC SDK 0.15.0 is live. This release adds multiple prompts batching, brings a native AMD GPU backend to the stack, moves more vision encoders onto mobile GPUs, and adds a second local coding-agent integration. Main highlights: - Prompt batching for the LLM addon. Batch multiple prompts into one job and process them concurrently, with each answer returned the moment its generation finishes. - Native AMD GPU backend. A first-class HIP/ROCm backend in @qvac/vla-ggml, auto-selected over Vulkan with clean fallback when ROCm is absent. - A second local coding agent. OpenClaw joins OpenCode for local, cloud-free agent workflows. AGENTS - OpenCode plugin update (@qvac/opencode-plugin). Aligned with the current SDK, CLI, and AI SDK provider packages. A fresh install runs OpenCode against managed local QVAC models out of the box, from the default qvac/qwen3.5-9b, with no manual qvac serve setup. - OpenClaw plugin (@qvac/openclaw-plugin). A second coding-agent integration alongside OpenCode. A fresh setup installs the plugin, creates a local qvac provider through onboarding, and runs a QVAC model through OpenClaw🦞's local service path. LANGUAGE MODELS - Prompt batching (LLM addon). Batch multiple prompts in one job and run them concurrently, each answer returns the moment its generation finishes, no waiting on the others. - Reasoning-context trimming on hybrid + recurrent models (@qvac/llm-llamacpp). remove_thinking_from_context now works beyond pure-attention models. Same JS API, no throw. VOICE AND SPEECH - Transcription (transcription-parakeet 0.9.0). More robust CPU fallback on GPU failure and a faster Vulkan backend on Pixel 9. - Text-to-speech features (tts-ggml 0.4.0). Adds LavaSR for noise removal and adjustable output frequency up to 48 kHz, plus Japanese via Chatterbox. - Text-to-speech fixes (tts-ggml 0.4.1). CPU fallback on GPU failure, a q8_0 KV crash fix on Metal with Chatterbox. VISION - Qwen3.5 vision encoder on GPU (Android). Image encoder moves onto the phone GPU, with a smarter tile-grid preprocessor and default image-token caps, for flagship Android: Vulkan on Mali (Pixel 9 Pro) and OpenCL on Adreno 830 (Galaxy S25). - Gemma-4 vision encoder on GPU (Android). Vision encoder runs on the phone GPU instead of CPU, same flagship Android targets. PLATFORM AND PERFORMANCE - AMD GPU backend (@qvac/vla-ggml). Native HIP/ROCm backend, auto-selected over Vulkan with clean fallback when ROCm is absent (Linux x64 only). Comes with ~23% faster than Vulkan, ~14% faster than PyTorch-ROCm, parity preserved. Unified code style. A cleaner, more consistent, easier-to-contribute codebase. Let's build. npm install @qvac/sdk

QVAC

17,700,244 views • 9 days ago

PhD Students – How to easily understand a complex research topic? Meet Ponder – a tool for understanding complex research. 𝐇𝐨𝐰 𝐏𝐨𝐧𝐝𝐞𝐫 𝐰𝐨𝐫𝐤𝐬? 1. Go to and log in 2. Enter your research topic or research question 3. Ponder will start building a knowledge map 4. This knowledge map breaks down complex ideas into structured cards 𝐖𝐡𝐚𝐭 𝐜𝐚𝐧 𝐲𝐨𝐮 𝐝𝐨 𝐰𝐢𝐭𝐡 𝐭𝐡𝐞𝐬𝐞 𝐜𝐚𝐫𝐝𝐬? → You can add your own thoughts, questions, and insights. → Ask follow-up questions and deepen your exploration. → You can color the cards for better understanding → You can drag & organize them freely across the infinite canvas. 𝐇𝐨𝐰 𝐭𝐨 𝐚𝐝𝐝 𝐫𝐞𝐬𝐞𝐚𝐫𝐜𝐡 𝐩𝐚𝐩𝐞𝐫𝐬 𝐭𝐨 𝐭𝐡𝐞 𝐜𝐚𝐫𝐝𝐬? — You can search for relevant papers with built-in discovery. — Ponder will identify all relevant papers — You can then add or upload research papers — You can also attach papers to specific cards. 𝐀𝐟𝐭𝐞𝐫 𝐲𝐨𝐮𝐫 𝐩𝐨𝐧𝐝𝐞𝐫𝐢𝐧𝐠 𝐢𝐬 𝐜𝐨𝐦𝐩𝐥𝐞𝐭𝐞𝐝: ➟ You can change the view to document, browser, or full screen. ➟ You can also download your knowledge map as a PDF ➟ You can ask further questions and refine with Ponder’s Agent. 𝐖𝐡𝐚𝐭 𝐮𝐧𝐝𝐞𝐫𝐬𝐭𝐚𝐧𝐝𝐢𝐧𝐠 𝐫𝐞𝐬𝐞𝐚𝐫𝐜𝐡 𝐭𝐡𝐢𝐬 𝐰𝐚𝐲 𝐢𝐬 𝟏𝟎𝐱 𝐛𝐞𝐭𝐭𝐞𝐫? ↳ It brings discovery and analysis of research into one workspace ↳ It makes ideas branch and evolve naturally, just like your brain ↳ It helps you to easily identify research gaps ↳ It connects knowledge from all sources such as papers and web ↳ It enables you to export knowledge as maps, reports, or data. ↳ Designed for PhD students & researchers, who think deeply. Try Ponder here: Anything you'd like to add?

Faheem Ullah

12,175 views • 1 year ago

QVAC SDK 0.14.0 is live. This release makes the on-device stack faster on mobile, ships the developer-agent path, and takes local text-to-speech to 31 languages. Main highlights: - OpenCode and OpenClaw. The first official OpenCode plugin, plus a maintained OpenClaw compatibility path, both built on managed mode and qvac serve. Point a coding agent at a local model with far less setup and far fewer surprises. - Brain-computer interface transcription, on the SDK. Take recorded neural signal data and decode it into text, fully on-device, no cloud. Stream it in chunks through a simple API. In 0.14 it runs GPU-accelerated on iOS. - Text to Speech in 31 languages with our Supertonic3 upgrade. VOICE AND SPEECH - Supertonic3 multilingual TTS, 5 languages to 31. - Chatterbox and Supertonic now run on the Android GPU, with lower memory use (especially on iOS), quantized s3gen Chatterbox support, and a fix for Chatterbox occasionally emitting random speech. - Whisper transcription now runs on the iOS GPU. Parakeet runs on the Android GPU, with steadier real-time streaming. VISION AND OCR - VLM multi-tile batching: high-resolution Pan and Scan images are encoded in one pass instead of tile by tile, for faster vision throughput. - OCR on ggml (EasyOCR and DocTR) reaches full speed parity with the onnx path, across Metal, OpenCL, and Vulkan. PLATFORM AND RELIABILITY - Dynamic compute backends on Linux: one build picks the right backend at runtime, and opens the door to ROCm and CUDA support without per-backend builds. - Thinking tokens are kept out of the model context, so reasoning no longer fills the KV cache. SDK 0.14.0 is now leaner and faster to start. Let’s build.

QVAC

23,973,950 views • 23 days ago