Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
yes's picture

yes

Gamination
1 1 31
ยท

AI & ML interests

None yet

Recent Activity

reacted to medmekk's post with โค๏ธ about 5 hours ago
๐Ÿš€ Introducing Halo 1.0 Today, we are open-sourcing Halo, the training framework we use to train every model at White Circle. It comes with: ๐Ÿง  Full post-training stack: SFT, DPO/KTO/SMPO, reward modeling, GRPO, distillation ๐Ÿค– Async multi-turn RL with vLLM/SGLang rollouts and sandboxed tool use โšก ~2.8ร— TRL throughput on 8ร— B300 (EP+FSDPv2, FA4, fp8/fp4) ๐Ÿค— Dense HF models + 15 MoE families (Qwen, GLM, Mistral, DeepSeek-V4โ€ฆ) ๐Ÿ› ๏ธ One halo command, prebuilt Docker images, and docs for humans and agents ๐Ÿ’ป https://github.com/whitecircle/halo Try it and tell us what you're training
liked a model about 22 hours ago
peculiar-ragdoll/Sharp-Spark-X2.5-4B-GGUF
liked a model 4 days ago
dealignai/Bonsai-2-27B-1bit-CRACK-GGUF
View all activity

Organizations

None yet

upvoted a collection about 2 months ago

Qwen 3.6 - Reg/Uncensored 9b, 12b, 21b, 27b, 40B

Collection
Fine tuned Qwen 3.6/3.8 models, including source and GGUF from 9B and up. 9B,12B, 21B and 40B are custom built by me. Tuning : Unsloth / COLD FUSION. โ€ข 26 items โ€ข Updated 5 days ago โ€ข 66
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs