qwen38-exl3-dflash2
qwen38-exl3-dflash2 is Qwen3.8-27B EXL3 (4.00 bpw) + DFlash2 speculative decoding for ExLlamaV3, validated at 262k context on a 24 GB RTX 3090 - r0b0tlab/qwen38-exl3-dflash2. It is ranked #258 on the Dev Radar, in Other, first seen 3 days ago and shared in 1 post (17.4k views).
GitHub - r0b0tlab/qwen38-exl3-dflash2: Qwen3.8-27B EXL3 (4.00 bpw) + DFlash2 speculative decoding for ExLlamaV3, validated at 262k context on a 24 GB RTX 3090 · GitHub Skip to content Navigation Menu Sign in Appearance settings Search / Sign in Sign up Appearance settings You signed in with another tab or window. Reload to refresh your session. You signed out in another tab or window. Reload to refresh your session. You switched accounts on another tab or window. Reload to refresh your session. Dismiss alert {{ message }} r0b0tlab / qwen38-exl3-dflash2 Public Notifications You must be signed in to change notification settings Fork 2 Star 17…
What people said about qwen38-exl3-dflash2 on X
r0b0tlab/Qwen3.8-27B-4.0bpw-EXL3 + DFlash2-EXL3 + optimized ExLlamaV3 runtime! Measured on one @NVIDIAAI RTX 3090 FE ⚡ 152.4 tok/s at 8K context 🧠 262K context capacity 💾 22.7 GiB peak VRAM runtime, weights, evals: https://github.com/r0b0tlab/qwen38-exl3-dflash2
— @mr_r0b0t, 3 days ago · 240 likes · see the post
Alternatives to qwen38-exl3-dflash2
- Ling-3.0-flash-Fin — Qwen-Drive-1.0-4B is a vision-language foundation model that handles 3D perception, driving VQA, and motion planning…
- bespoke-nimble-9b — We’re on a journey to advance and democratize artificial intelligence through open source and open science.
- Cactus Compute — It runs on mobiles, wearables, smart home devices, small robots and microcontrollers, with prebuilt engines for macOS,…
- orcabonsai-27b-uncensored — Open source
- Spark-X2.5 — 🤖 ModelScope
- grok.com —
qwen38-exl3-dflash2 in numbers
- Rank on the Dev Radar: #258 of 1355
- Shared in 1 post by 1 account: @mr_r0b0t
- 17.4k views on those posts
- First seen 3 days ago, last shared 3 days ago
- Pricing seen by Jev: open source
- Market: Other
FAQ
What is qwen38-exl3-dflash2?
Qwen3.8-27B EXL3 (4.00 bpw) + DFlash2 speculative decoding for ExLlamaV3, validated at 262k context on a 24 GB RTX 3090 - r0b0tlab/qwen38-exl3-dflash2 It was first shared on X 3 days ago and is ranked #258 on the Dev Radar.
Is qwen38-exl3-dflash2 free?
It is open source.
Who shared qwen38-exl3-dflash2?
1 account on X, including @mr_r0b0t, in 1 post totalling 17.4k views.
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 12.2k posts from 4.7k X accounts over the last 21 days, 1.4k tools, 12 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:39 UTC. Full method.