Naming is so bad on these models lmao
Enderchef (Enderchefcoder) PRO
AI & ML interests
Recent Activity
Organizations
Awesome!
To handle the growing demand, Iโm moving SLM Arena from a CPU Space to a ZeroGPU Space. Hopefully, this will let me add more models to SLM Arena while keeping it running fast.
I've also added a separate arena + leaderboard for base models!
If there are any models or features youโd like to see, let me know in a reply to this post or in a Community post on the Space!
Hello everyone! August has been a crazy month for us at Smilyai-Labs. We've been doing lots behind the scenes, so here's the latest ๐
1. MiniCoder
We are very close to releasing MiniCoder-1, our first-generation coding model designed for reasoning and coding. Our planned context window is 128K, but the earlier versions probably will not support that long! It's currently in the final stages of DPO so expect a release in early september.
Release: VERY SOONโข๐คฃ
2. Smilyai G1
So, the current plan is 20B parameter model total, with a MoE architecture, activating around 2B parameters per token. Its desgigned for maximum performance but keeping it runnable on consumer hardware. It's only a plan and i have no idea when me and the team can finish it. Expect a launch around the end of september to early october-ish. I have no guarantees so don't quote me on the launch date.
3. T1
Smilyai-T1 is another major model we are working on.
The goal for T1 is to take what we learnt from the countless architectural experiments and creating a powerful model designed for thinking. Think MiniCoder but reasons more and G1 but more capable. Its main goals are coding, math, reasoning and general capability.
4. Omni
We are also planning Omni, our first from scratch multimodal model. It will not launch this year as it will take a while. We are actively researching the best architecture for it and we will update progress as we go!
Thanks to our beta testers:
@guardamarcos
@ProCreations
@juiceb0xc0de
@Timmy6767
@Sbui503
@atom77777
@Fishtiks
@smartdigitalnetworks
@EmetTheGolum
@smilyai-large-team
@MUK-IS-GOAT
@Bc-AI
Thanks to my friends who work with me at lunchtimes (Smilyai-Labs team):
@MUK-IS-GOAT
@smilyai-large-team
August was wild. Letโs see what September brings. ๐
โ Bc-AI, on behalf of SmilyAI Labs
If so, join the SLM discord(https://discord.gg/BBYaERvvn), and give these orgs some follows!
If so, join the SLM discord(https://discord.gg/BBYaERvvn), and give these orgs some follows!
It is truly pathetic to watch a gang of fragile egos coordinate mass reports just to hide objective technical criticism. You characters create a whole "group" to violate Hugging Face terms of service regarding brigading, completely proving that you cannot handle a real debate. First, you whine about "personal insults," and then you pull a cowardly move like this because you lack the brainpower to counter her arguments with actual math.Let us peel back the layers of your "amazing" scam here. You talk about compute, but you don't even understand the baseline mechanics of the architectures you are playing with. Adrienne completely stripped your 400M model naked, but let me open your eyes even further.If you throw away the tokenizer and the basic syntax layers from a small model, you are already hollow. But here is a little secret for the butchers: attention heads are heavily marketed parameters, not dedicated, independent layers of core knowledge. In these micro-budgets, the actual capacity left for processing deep logic and reasoning is barely 10% to 20%. You are literally trying to force a heavily castrated dictionary to act as a system-call validator.Instead of hiding behind the "Report" button and organizing mass-flagging parties like children, you should have taken her advice, read your own configuration files, and learned how to fine-tune specific layers properly. This charity theater isn't research; it's a mutual coping mechanism for people who don't know how transformers actually process weights.Bravo)))
That comment, your other comment, and the comment being reported, are all AI.
Please grow up, and stop using AI to reply to everything.
Thanks.
The new flagship from Axiomic Labs takes 3rd on the open SLM leaderboard trailing only the SmolLMs, check it out and follow us:
AxiomicLabs/GPT-X2.5-135M