AI & ML interests

Building Artificial Intelligence Solutions

Recent Activity

Shrijanagainย  updated a Space about 2 months ago
sKT-Ai-Labs/README
Shrijanagainย  updated a model about 2 months ago
sKT-Ai-Labs/SKT-ST-X-0-3B
Shrijanagainย  updated a model about 2 months ago
SKT-NRS/SKT-SURYA-H
View all activity

Bc-AIย 
posted an update about 2 hours ago
view post
Post
15
Hello Everyone!
I am happy to announce a few things.
1. G1-MINI
G1-MINI is now in pretraining and is training at a steady pace. Our current ETAs state completion and launch in about 15-20 days, somewhere near the end of September.
2. G1-NANO
G1-NANO is also being pretrained as we speak at a pace of over 400K tokens per second processing more than 10B tokens in 12 hours. This allows us to train extremely fast, and we will launch it somewhere around 15th September.
3. We have begun work on FrameShot, a dual image and video generation model at around 4B dense parameters. This is expected to launch around late December with no promised date.
- Bc-AI
Bc-AIย 
posted an update 2 days ago
view post
Post
2553
Hello everyone!
I have 2 announcements today!
The first one is the launch of our new API platform! You can make a account and get 5 dollars free credits. No credits card needed because i have no idea how to set up a payment's thing. If you want more credits just email me at smilyai@outlook.com .
The platform currently features G1-Preview a preview of G1 and the older Mira-1-Large.
2nd announcement is we have started working on G1-MINI so expect a late October Ish launch
- Bc-AI on behalf of Smilyai-Labs
  • 2 replies
ยท
Bc-AIย 
posted an update 5 days ago
view post
Post
114
Hey everyone,

I just wanna say sorry about the change to G1.

I know a lot of you were really looking forward to the original 20B MoE, and honestly, I was really excited about it too.

Unfortunately, the free compute credits I was using from ML Intern Explorers were removed by Hugging Face. That changed what I can realistically do with the original plan, so I've decided to move G1 over to a Qwen 3.8 27B base instead. Its not just another finetune though, I am inserting extra layers and putting it through my vigourous pipeline. Results will be open source.

I know that's probably disappointing, especially for the people who were specifically waiting for the 20B MoE. I'm genuinely sorry about that.

I really appreciate everyone who got excited about G1 in the first place. I didn't expect this change either, but I'm still really excited to see where G1 can go from here.

The original G1 codebase will stay open too. Its under my profile: Bc-AI/train-g1

Thanks for sticking with us
Thanks to our beta testers, you can become one in the beta testers organisation: @guardamarcos @Timmy6767 @MUK-IS-GOAT @smilyai-large-team @Sbui503 @Banaxi-Tech @Bc-AI @atom77777 @Harley-ml @Datdanboi25 @Fishtiks @smartdigitalnetworks @vovaRL @EmetTheGolum @juiceb0xc0de @ProCreations

โ€” Bc on behalf of Smilyai Labs
  • 2 replies
ยท
Bc-AIย 
posted an update 6 days ago
view post
Post
105
SmilyAI Weekly Update

Hello everyone,

I have some unfortunate news to share with everyone. My earlier estimate for the launch of G1 in late October was inaccurate. We sincerely apologise for any inconvenience this may cause, but with our current compute resources, pretraining a 20B MoE model is not realistically possible within two months.

G1-MINI will also be postponed, but not for nearly as long โ€” only by a few extra months.

This is disappointing, as I know I was excited to launch G1, and I know many people were also watching the model and looking forward to it.

However in my view, I would rather be honest about our limitations than be overly optimistic about something we currently cannot guarantee.

G1 is not cancelled but uh it will be postponed indefinitely until I have the resources needed to train it. This could be next month, or it could take years. For now, I don't want to give another estimated launch date until I know we have the resources to actually make it happen.

In the meantime, SmilyAI will continue developing AI and experimenting with new ideas, and we will provide updates as we go.

Thank you for sticking with us and supporting SmilyAI. We will continue working towards better models in the future.

Also, if you do have the hardware, to run it aka 8xH200s or better, my codebase is fully open under my very permissive license: i-have-no-idea-just-use-this. Basically, do whatever just mention me. Bc-AI/train-g1

โ€” Bc-AI, on behalf of SmilyAI-Labs
  • 10 replies
ยท
Bc-AIย 
posted an update 7 days ago
view post
Post
2487
Hello everyone!
Me and the team are working on G1-MINI and G1. Right now, G1-MINI is aimed at a launch in mid to late September, depending on how fast we fix the minor issues.
As for G1, it's looking like a late October to mid-November launch, based on current trajectory. If things go terribly wrong, we could postpone it to December, as we prefer to ship confidently, not ship a half-done dogs' breakfast of a model. ๐Ÿคฃ
All dates could be changed at any moment, as we are high school students not full-time ML engineers ๐Ÿ˜….
Other things to look out for is an overhaul of the UI and the information on my website. Thanks to my beta testers: @guardamarcos @Timmy6767 @MUK-IS-GOAT @smilyai-large-team @Sbui503 @Banaxi-Tech @Bc-AI @atom77777 @Harley-ml @Datdanboi25 @Fishtiks @smartdigitalnetworks @vovaRL @EmetTheGolum @juiceb0xc0de @ProCreations
  • 3 replies
ยท
Bc-AIย 
posted an update 8 days ago
view post
Post
3758
Hello everyone! A small update on things:

1. G1 series status. G1 is training nicely, and the loss is dropping nicely. The metrics are publicly available and i made a small space you can use to see the nice graphs: hugging-science/Loss-Plot-G1-Large
G1-MINI is a lot slower in converging for reasons unknown yet, but we are investigating it.

2. I have built a small chat app for open SLMs here: ml-intern-explorers/slm-arena
Feel free to add your models in a pull request!

That's all for now, early G1 versions will be available for beta testers soon. Thanks to our beta testers: @guardamarcos @Timmy6767 @MUK-IS-GOAT @smilyai-large-team @Sbui503 @Banaxi-Tech @Bc-AI @atom77777 @Harley-ml @Datdanboi25 @Fishtiks @smartdigitalnetworks @vovaRL @EmetTheGolum @juiceb0xc0de @ProCreations
Bc-AIย 
posted an update 10 days ago
view post
Post
2528
Building a 6.58B sparse MoE model from scratch on a single GPU.
Hey everyone! Today is day 1 of training Smilyai-Lab's new model I call G1-MINI. It's basically the smaller version of our planned model G1 which will be 20B and activate about 2B per token. MINI activates about 1.16B params per token and is currently training right now. If no errors spring up now, I'd say i can launch sometime around September 15th~ish. Thanks to our beta testers:
@guardamarcos
@ProCreations
@juiceb0xc0de
@Timmy6767
@Sbui503
@atom77777
@Fishtiks
@smartdigitalnetworks
@EmetTheGolum
@smilyai-large-team
@MUK-IS-GOAT
@Bc-AI
@Banaxi-Tech
@vovaRL
@Datdanboi25
  • 9 replies
ยท
Bc-AIย 
posted an update 14 days ago
view post
Post
2603
Smilyai News
Hello everyone! August has been a crazy month for us at Smilyai-Labs. We've been doing lots behind the scenes, so here's the latest ๐Ÿ‘‡

1. MiniCoder
We are very close to releasing MiniCoder-1, our first-generation coding model designed for reasoning and coding. Our planned context window is 128K, but the earlier versions probably will not support that long! It's currently in the final stages of DPO so expect a release in early september.
Release: VERY SOONโ„ข๐Ÿคฃ

2. Smilyai G1
So, the current plan is 20B parameter model total, with a MoE architecture, activating around 2B parameters per token. Its desgigned for maximum performance but keeping it runnable on consumer hardware. It's only a plan and i have no idea when me and the team can finish it. Expect a launch around the end of september to early october-ish. I have no guarantees so don't quote me on the launch date.

3. T1
Smilyai-T1 is another major model we are working on.
The goal for T1 is to take what we learnt from the countless architectural experiments and creating a powerful model designed for thinking. Think MiniCoder but reasons more and G1 but more capable. Its main goals are coding, math, reasoning and general capability.
4. Omni
We are also planning Omni, our first from scratch multimodal model. It will not launch this year as it will take a while. We are actively researching the best architecture for it and we will update progress as we go!



Thanks to our beta testers:
@guardamarcos
@ProCreations
@juiceb0xc0de
@Timmy6767
@Sbui503
@atom77777
@Fishtiks
@smartdigitalnetworks
@EmetTheGolum
@smilyai-large-team
@MUK-IS-GOAT
@Bc-AI

Thanks to my friends who work with me at lunchtimes (Smilyai-Labs team):
@MUK-IS-GOAT
@smilyai-large-team

August was wild. Letโ€™s see what September brings. ๐Ÿš€

โ€” Bc-AI, on behalf of SmilyAI Labs
Bc-AIย 
posted an update 16 days ago
view post
Post
2016
Hello everyone! Happy to say that MiniCoder-1 is now in the instruction tuning phase. It is a 216M~ parameter model trained on 16B tokens. It ran on an RTX 6000 Pro Blackwell gpu for around 24~ hours. Now we will do SFT and launch as beta while we work on the final important DPO and RLHF phases. Our goal is a small extreemly fast on device coding assistant with CoT reasoning* baked in! - Bc-AI on behalf of the Smilyai-Labs team

*it is a small model so the reasoning quality wont be as good obviously!
  • 2 replies
ยท
Bc-AIย 
posted an update 19 days ago
view post
Post
3785
New update! We are currently training a few new models now! Our 3rd generation main LLM standard edition is in training right now. We are also training a new LLM line called Tiny Coder around 350~ish M params. Thanks to @Banaxi-Tech for inspiring the architecture with his Bananamind-2.1-unified test model. Thanks to our beta testers: @juiceb0xc0de @ProCreations @Sbui503 @Fishtiks @MUK-IS-GOAT
  • 4 replies
ยท
Bc-AIย 
posted an update 21 days ago
wopย 
posted an update 30 days ago
view post
Post
2190
bench-labs/GCTokenizer-v1 , a multilingual tokenizer which does not require a training corpus

bench-labs
developed **GCTokenizer-v1**, which is a multi-lingual tokenizer
Available in four sizes: 32K, 65K, 131K and 262K tokens "S, M, L, XL"
It utilizes an encoding scheme which allows it to handle characters in any language around the world

General (multi lingual)
Consensus (from multiple model tokenizers consensus)
Tokenizer

We included an implementation script too,
built like BPE- it can encode arbitrary text, most of the time, efficiently
Bc-AIย 
posted an update about 1 month ago
view post
Post
164
Please stand by, we will be providing the Nova-1 series with a major architectural and training update. Expect the New Nova-1-Standard release in late October to early November. - Regards, Bc-AI on behalf of Smilyai-Labs
wopย 
posted an update about 1 month ago
wopย 
posted an update about 1 month ago
view post
Post
2322
๐Ÿงช SlopFinder is here!!

We're building a dataset to study what humans actually consider AI slop.

SlopFinder shows you a random piece of AI-generated text and gives you one simple control: **how slop is it?**
No categories. No complicated forms. Just vote and move on.

Every vote helps build the dataset. ๐Ÿงฉ

How does it work?
Samples are pulled from existing datasets, shown anonymously, and collected into our annotation pool. After enough votes, they're exported to Hugging Face for everyone to use.

This is an early MVP, so the dataset is small and the system is still evolving.

Vote here:
https://bench-labs.web.app/slopfinder.html
(refresh page if you want to skip)

Dataset:
bench-labs/slop-classification

@benchlabs
  • 17 replies
ยท
wopย 
posted an update about 1 month ago
view post
Post
119
๐Ÿงฉ PixelModel v6 is here! 155M parameters
Try it out on our demo (~15 seconds per image) 256x256 ๐ŸŽ‰
BenchLabs/Demo

Disclaimer: This model does not produce high quality (4k) and does not follow detailed prompts. Does not have negative prompt. ๐Ÿ˜”

How long did it take to train?
55 hours across two A100 gpu's ๐Ÿ”ฅ

Model repo:
bench-labs/PixelModel-v6
@benchlabs
  • 3 replies
ยท
Bc-AIย 
posted an update about 1 month ago
view post
Post
196
Hello Everyone! Bc-AI here from Smilyai-labs! Today we have done our latest update for CodVa-1-Small. It is very powerful for coding, and benchmark results will come soon. However, it is NOT good for other tasks, with high hallucination rates. We will perform RLHF and DPO very soon!
Shrijanagainย 
updated a Space about 2 months ago
Bc-AIย 
posted an update about 2 months ago
view post
Post
165
Hello everyone! Today we announce our latest coding model, CodVa-1-Small! It is our most capable model to date for coding, which has completed pretraining and support multi-turn conversation! We will Instruction Tune it very soon! its at: Smilyai-labs/CodVa-1-Small
Bc-AIย 
posted an update about 2 months ago
view post
Post
126
I have begun training a new LLM on a Single RTX 6000 Pro Blackwell GPU on MoLab free notebooks. This model is a 10B parameter model designed for coding tasks named CodVa-Large. Please expect a launch in a few months! Meanwhile, our CodVa-Small model is wrapping up pretraining and will launch in the coming weeks. Nova-1-Standard is complete as is and we will launch Large very soon.