DFlash 2: Qwen3.8-27B at 2× Speed - Live Benchmark Locally
Local installation and testing of Qwen3.8-27B using Dflash2, featuring live local benchmarking performance evaluations.
Videos
Find videos about LLM models, hardware, runtimes, benchmarks and practical AI workflows.
Local installation and testing of Qwen3.8-27B using Dflash2, featuring live local benchmarking performance evaluations.
Demonstrates Qwen3.8-27B running unsupervised local coding tasks for 10 hours, including building a Pi/goal extension via Grok Build and optimizing SGLang-Omni to serve MiniMax-Music3.
Installs and tests hermes agent loop functionality using Qwen3.8 27B to demonstrate recurring execution of prompts or slash commands.
Evaluates Grok 4.6 and Grok Bot against Fable 5, GPT 5.6, and Hermes Agent via BridgeMind 1 tasks, fal.ai plugins, Remotion teasers, and crash fixes using speed and cost metrics.
Explores TalkWithMe v4 features including voice cloning, persona editor, echo chamber and true group chat using dots.tts or Qwen3-TTS.
Demonstrates SystemMessage, HumanMessage, and AIMessage mapping, LangSmith tracing setup, pirate/support/rhyme prompt tests, token cost measurement, and prompt injection against a no-refund agent usin
Demonstrates generating test cases from requirements, authoring and executing scripts, debugging failures, and running on 30,000+ real devices via Playwright, Selenium, and Cypress within the IDE.
Installs and tests Qwen3.8-27B-Ridge-3.7bpw using GGUF format from empero-ai repository.
Compares Qwen3.8 27B and Qwen3.6 27B using Regina on Minesweeper, VMs, py2c calc, encryption and memcached, analyzing reasoning tokens, degeneracy and failures.
Evaluates Qwen 3.8-27B quantizations (BF16, Q8, Q4 Unsloth, Q4_K_M Ollama, Q2) via one-shot exams measuring speed and quality against a BF16 baseline.
Installs free open source LTX 2.5 in ComfyUI, demonstrating text to video, image to video, first last frame generation, using loras, low VRAM settings, Minimax H3 vs LTX 2.5 comparison, Minimax plus L
Tests Qwen 3.8 using Inferencer App v2.3.3 on M3 Ultra 512 GiB across 2 million tokens, varying temperature, quantization levels, and thinking modes.
Installs and tests Antares-1B for vulnerability localization in real-world codebases.
Evaluates Qwen 3.8 27B using an eight-stage development plan on improved hardware achieving 70–130 tokens per second, with code review included.
Evaluates Qwen3.8-27B inference via llama.cpp RPC across RTX 3060 and Apple Silicon Mac, comparing throughput metrics for single-card versus distributed layer loading.
Tests Gemini 3.7 Flash on browser workflows, 3D scenes, C++ games, CAD models, hardware mods, FPS dev, frontend design, and the Street Yeet game.
Demonstrates making a brownfield codebase AI native across PM, developer, and QA roles using the AI Layer Starter Pack, global rules, and automated PR reviews in CI.
Tests DeepSeek V4 Pro (0813) via WoAI Bench on frontend dev, agentic coding, 3D/Three.js, and one-shot gen, comparing cost, speed, and quality against Gemini 3.7 Flash, Grok 4.6, Kimi K3, GLM 5.2, Qwe
Installs and tests Qwen3.8-9B, described as a full-parameter distillation of Qwen3.8 2.4T A95B.
Qwen 3.8-27B versus Qwen 3.6-35B-A3B at BF16 on a custom exam featuring Tetris, volcano physics, spreadsheets, and SVG tasks, plus Claude Opus 4.6 comparison.
Reviews GLM 5.3 capabilities through demonstrations involving Z Code, Windows replica creation, Blender V8 engine usage, 3D fighting games, music composition, financial presentations, deep research, c
Evaluates DeepSeek V4 Pro and V4 Flash via BridgeMind horror game, Minecraft clone, Next.js dashboard, and BridgeMind 1 bug against GPT 5.6, Fable 5, and Grok 4.6 using DeepSeek Harness.
An app based on local AI benchmarks assists users in determining suitable hardware configurations for running specific machine learning models locally.
Installs and tests LTX-2.5, described as an open world model, within ComfyUI, covering setup fixes and initial generations using resources from huggingface.co/Lightricks/LTX-2.5.
Configures Claude Code to use Nemotron 3.5 Lightning via Ollama on an RTX 5090 through .claude/settings.local.json, verified by generating and modifying a Python palindrome function.
Tests GLM 5.3 on browser workflows, C++ game development, 3D CAD modeling, frontend design, FPS generation, and a difficult hardware test.
Demonstrates installing Unsloth Studio via terminal on AMD ROCm hardware, loading the full 55 GB Qwen 3.8-27B model consuming 83 GB RAM, and testing local chat, vision, and deep research capabilities.
Evaluates Qwen 3.8 27B against Qwen 3.6 27B on psim-blackbox using pi harness 0.79.3 across low, medium, and xhigh reasoning settings, showing performance varies significantly by mode.
Overview of Deepseek V4 0813, Qwen 3.8 variants, GLM 5.3, Grok 4.6, LTX 2.5, Gemini 3.7 Flash, GPT Ultrafast, JoyAI Video Edit, Scope, Index TTS 2.5, Minimax Music 3, Magi 2, Nemotron Lightning, and o
Locally installs and tests DeepSeek Harness (dsh), an open-source agent harness.
Evaluates Qwen 3.8-27B, a native multimodal dense model reading images and video, covering Unsloth quants, thinking-effort controls, NVFP4 Blackwell builds, and benchmarks against Opus 4.6 Max.
Side-by-side coding test using Claude Code on Ollama to generate Space Invaders, Breakout, and Tetris HTML canvas games with Qwen 3.8 27B, Muse Glimmer, and Gemma 4 on an RTX 5090.
Tests Qwen3.8-27B via live tool calling with Playwright MCP Server and migrating a .NET solution from .NET 8 to .NET 10, while examining Preserve Thinking for multi-turn agentic work.
Tests how many AI agents the ASUS ExpertCenter Pro ET900N G3 can run using its 748GB unified memory, 400Gb networking, and 1400 watt superchip configuration.
Tests unsloth/Qwen3.8-27B-GGUF performance, memory, agency, OpenAI Human Eval, Kanban, Sand Physics, Dungeon Crawler, Blender and Godot on a 16GB local setup using llama.cpp.
Evaluates Qwen 3.8 27B locally through code reviews of a nontrivial feature branch and analysis of technical documentation for gaps or outdated information.
Tests Qwen 3.8 27B, Nemotron 3.5 Lightning, and Muse Glimmer on an RTX 5090 using 16 hard math, coding, and reasoning problems via Ollama with Q4_K_M quantization, measuring accuracy, time, VRAM, and
Tests Qwen3.8-27B locally on RTX 4090 via Unsloth Dynamic 4-bit quant and Open WebUI against Claude Opus in coding, agentic tasks, vision and web dev.
Installs and tests Qwen3.8-27B Quant locally using llama.cpp, Ollama, and LM Studio.
Evaluates Qwen3.8-27B-MLX-Q9 performance using Inferencer App v2.3.2 on M3 Ultra 512 GiB across one million tokens.
Demonstrates setup of Qwen 3.8 27B bootstrapped by Qwen 3.6 27B on local infrastructure, where Qwen 3.8 27B programs Super Dario.
Tutorial demonstrating Minimax Music 3 installation and workflows in ComfyUI, covering pop, punk, jazz genres, song editing, loop stacking, and conversion to MIDI.
Tests Qwen3.8 27B across browser OS, C++ skate game, Subway FPS, 3D CAD, watch website, creative chat, multimodal coding, Chrono City timeline, Cinematic Steve’s PC repair, and Street Yeet game.
Demonstrates building an AI Dark Factory using Archon and a ColeAm00 skill to automate specs into validated software via five components: workflow-driven repos, automation queues, blue-green deploymen
Locally evaluates Qwen3.8-27B against Qwen3.6-27B on coding, agents, tools, reasoning, hallucinations, speed, vision, Splunk CTF, and Mario clone creation.
Qwen3.8-27B runs locally on MacBook Pro M5 Max 128GB, evaluating speed, reasoning, coding, and offline performance via Pi coding agent and Mario game construction tasks.
Installs and tests Qwen3.8-27B locally.
Covers WorldClaw, Grok Bot, Seedance 2.5, Claude labels, Suno limits, Spotify badges, Grok 4.6, Gemini 3.7 Flash, Meta Muse Glimmer, NVIDIA Nemotron 3.5 Lightning, DeepSeek-V4-Pro, MAI-Code-1.1-Flash,
Tests GLM-5.3 on coding, frontend, agentic tasks, Terminal Bench, DeepSWE, and AutomationBench, comparing against Kimi K3, DeepSeek-V4, Qwen3.8-Max, GPT-5.6 Sol, Opus 4.8, and Fable 5.
Tests local installation of MiniMax-Music3 for generating Bollywood, Cumbia and pop music using resources from huggingface.co/MiniMaxAI/MiniMax-Music3.