model roundup

Qwen 3

19 items · started 2026-04-25 · closed 2026-05-09

  1. WebWorld is a large-scale open-web world model series for training and evaluating web agents. It is trained on 1M+ real-world web interaction trajectories via a scalable hierarchical data pipeline, supporting: Long-horizon simulation (30+…

  2. Wanted to see if a real voice loop — speak, model thinks, speaks back — could run entirely on a single device today, no cloud. Same codebase doubles as a live translator (speak in language A, hear it back in language B).

  3. emotion-steering Extract and serve CAA-style emotion steering vectors for any HuggingFace causal LM, with a fast vLLM path for Qwen3. ┌────────────┐ ┌────────────────────┐ labeled │ extract │ vectors + AUC report │ serve │ contrasts ├─────…

  4. Hello guys, So TL;DR, I was asked by multiple people to make an Assistant_Pepe_32B version, but the best base model contender was Qwen3-32B, a model that is very hard to tune on anything other than STEM. The concept of Assistant_Pepe is an…

  5. dl.acm.org Performing security verification This website uses a security service to protect against malicious bots. This page is displayed while the website verifies you are not a bot.

  6. I have tried Qwen 3 VL family of models on my rtx3060, max I can load is Q8 8b. The task is visual reasoning/ instruction following.

  7. Talk at Qwen Meetup Korea end of May. Looking for review on this draft before I build PPT slides off it.

  8. Got a chance to check this model today. 8GB VRAM(RTX 4060 Laptop GPU) & 32GB DDR5 RAM.

  9. Hi everyone, wanted to see how far QVAC could be pushed on a phone: full speech-to-text → LLM → text to-speech running locally, no network, and get it close to a real conversation. Stack (Android, all via qvac sdk): - STT: Parakeet (stream…

  10. https://preview.redd.it/7yei65sbugyg1.png?width=1703&format=png&auto=webp&s=ad388c51dd10cb44b41a99876d28797e006fd138 Stanford's Generative Agents = one LLM cosplaying 25 personas. I wanted agents that actually become different people — dif…

  11. Just got myself (for my company) a RTX Pro 6000 Blackwell Workstation card. Managed to get really good TPS on qwen3 27b fp8.

  12. Hey everyone, I’ve been working on Chirp, a native offline text-to-speech desktop app. It runs locally on your machine, supports both Kokoro and Qwen3-TTS, and is written in C++ and Rust.

  13. I've been benchmarking open source omni models like Qwen3-Omni for speech to speech tasks and they perform... really well.

  14. I'm genuinely surprised at what people are willing to share with AI companions. Read Replika's privacy policy.

  15. I just got Qwen3 72B Instruct running on a high RAM setup and I’m kinda confused about the proper way to use it. What’s the correct workflow for running it smoothly (like best quant, tools, or runtime)?

  16. Going through university right now, and we have massive 100 page pdfs/ppts with soo much fluff that its annoying to go through. until now ive been using chatgpt for it, but realized that the output tokens are HEAVILY limited, and loses a L…

  17. If you have low vram - qwen 3 tts is good If you need something unique go for - tada 3b but it need 28gb vram If you want best tts rn + have the commercial use allowed then go for - moss tts 8b its literally the best model out there Litera…

  18. Heya guys and gals, Around a year ago I released and posted about Persona Engine as a fun side project, trying to get the whole ASR -> LLM -> TTS pipeline going fully locally while having a realtime avatar that is lip-synced (think VTuber)…

  19. All I see is "it gives me **/s bla bla bla" all together with q4, q3... even when chatting with qwen3.

← all threads