view post Post 2731 Qwen-Image-2.1 Plug and Play LoRA App is now live on Hugging Face Spaces.🔗 Space: prithivMLmods/Qwen-Image-2.1-LoRAs-PnPIt supports standard inference, 4-step Turbo inference, custom LoRA lazy repacks, and LoRA Plug and Play (PnP), all in one setting! 🔗 Qwen-Image-2.1 Image-to-Image LoRAs: https://huggingface.co/collections/prithivMLmods/qwen-image-21-image-to-image-loras🔗 GitHub: https://github.com/PRITHIVSAKTHIUR/Qwen-Image-2.1-LoRAs-PnPTo learn more, visit the app page or the respective model pages. See translation 🔥 3 3 ❤️ 3 3 👍 2 2 🚀 2 2 🤗 1 1 + Reply
Realtime-Venus: A full-duplex interaction system with asynchronous delegation Paper • 2609.13814 • Published 16 days ago • 222
view post Post 4314 🚀 Introducing Halo 1.0Today, we are open-sourcing Halo, the training framework we use to train every model at White Circle.It comes with: 🧠 Full post-training stack: SFT, DPO/KTO/SMPO, reward modeling, GRPO, distillation🤖 Async multi-turn RL with vLLM/SGLang rollouts and sandboxed tool use⚡ ~2.8× TRL throughput on 8× B300 (EP+FSDPv2, FA4, fp8/fp4)🤗 Dense HF models + 15 MoE families (Qwen, GLM, Mistral, DeepSeek-V4…)🛠️ One halo command, prebuilt Docker images, and docs for humans and agents💻 https://github.com/whitecircle/haloTry it and tell us what you're training See translation 1 reply · 🚀 9 9 ❤️ 4 4 + Reply
Training Specialist Models without Reasoning Trajectories for Domain Expert Distillation Paper • 2609.13770 • Published 16 days ago • 9