Join the conversation

Join the community of Machine Learners and AI enthusiasts.

Sign Up

All HF Hub posts

OppaAIΒ 
posted an update 2 days ago
appvoidΒ 
posted an update 2 days ago
view post
Post
2262
Nobody knows what is doing, when you train a model, you are experimenting to advance the frontier, so keep failing 🫡
  • 12 replies
Β·
SeaWolf-AIΒ 
posted an update about 13 hours ago
view post
Post
529
πŸ§ͺ Open Discovery Challenge β€” Season 4 is open: non-opioid pain
WHO titled its 2023 report "Left behind in pain."

The same drug kills by excess in one part of the world and, by its absence, lets people die in agony elsewhere. About 80% of the ~600,000 drug-related deaths WHO estimated for 2019 involved opioids. The same report records a 5-fold to 63-fold gap in morphine consumption between rich and poor countries: the richest 10% use 90% of what circulates. Everyone else endures surgery, and terminal cancer, without it.

Both problems have one answer: a painkiller that does not create dependence.

Nav1.7 has come closest. People born without a working copy of this channel feel no pain while every other sensation stays normal β€” validated not in animals but in humans.

There is still no drug, and the difficulty is not the target but the discrimination. The body carries several similar sodium channels, and blocking the heart's hERG channel alongside causes fatal arrhythmia. Several candidates were discontinued for exactly that.

Season 4 asks one question: can you block the pain channel alone?

Target β€” Nav1.7 VSD4, the domain IV voltage sensor where this inhibitor class binds
Anti-target β€” hERG pore, computed as the tetramer: four subunits together form the space a drug enters, and a monomer misses the binders that matter.
Closes 2027-01-31 Β· Prize USD 1,000 to the season's #1
Any model, any harness. However you found the candidate, it meets the same rubric.

14 days, 9,886 candidates, 108 participants
ODC opened on 2026-08-15. In the fourteen days since, 9,886 candidate molecules have come from 108 participants across four seasons β€” malaria, tuberculosis, Chagas disease, and now non-opioid pain. About 700 a day, from people who mostly do not know each other.

The candidates are the point. The leaderboard is only how we keep score.

πŸ‘‰ FINAL-Bench/open-discovery-challenge
CodeSoftΒ 
posted an update 1 day ago
view post
Post
1957
Wow, SLM Arena is getting a lot of traffic! Thank you guys for showing your interest!

To handle the growing demand, I’m moving SLM Arena from a CPU Space to a ZeroGPU Space. Hopefully, this will let me add more models to SLM Arena while keeping it running fast.

I've also added a separate arena + leaderboard for base models!

If there are any models or features you’d like to see, let me know in a reply to this post or in a Community post on the Space!
  • 17 replies
Β·
Bc-AIΒ 
posted an update 3 days ago
view post
Post
2506
Smilyai News
Hello everyone! August has been a crazy month for us at Smilyai-Labs. We've been doing lots behind the scenes, so here's the latest πŸ‘‡

1. MiniCoder
We are very close to releasing MiniCoder-1, our first-generation coding model designed for reasoning and coding. Our planned context window is 128K, but the earlier versions probably will not support that long! It's currently in the final stages of DPO so expect a release in early september.
Release: VERY SOONβ„’πŸ€£

2. Smilyai G1
So, the current plan is 20B parameter model total, with a MoE architecture, activating around 2B parameters per token. Its desgigned for maximum performance but keeping it runnable on consumer hardware. It's only a plan and i have no idea when me and the team can finish it. Expect a launch around the end of september to early october-ish. I have no guarantees so don't quote me on the launch date.

3. T1
Smilyai-T1 is another major model we are working on.
The goal for T1 is to take what we learnt from the countless architectural experiments and creating a powerful model designed for thinking. Think MiniCoder but reasons more and G1 but more capable. Its main goals are coding, math, reasoning and general capability.
4. Omni
We are also planning Omni, our first from scratch multimodal model. It will not launch this year as it will take a while. We are actively researching the best architecture for it and we will update progress as we go!



Thanks to our beta testers:
@guardamarcos
@ProCreations
@juiceb0xc0de
@Timmy6767
@Sbui503
@atom77777
@Fishtiks
@smartdigitalnetworks
@EmetTheGolum
@smilyai-large-team
@MUK-IS-GOAT
@Bc-AI

Thanks to my friends who work with me at lunchtimes (Smilyai-Labs team):
@MUK-IS-GOAT
@smilyai-large-team

August was wild. Let’s see what September brings. πŸš€

β€” Bc-AI, on behalf of SmilyAI Labs
Hoglet-33Β 
posted an update 3 days ago
view post
Post
2938
We are announcing the first generation of the Pebble model family!

These are the models we are releasing:

- Pebble 10M
- Pebble 25M
- Pebble 50M

Each model will use a Mamba-Transformer 3:1 hybrid architecture and will be pretrained on 25 billion tokens before IFT and SFT.

Depending on development time and resources, we may also release:

- Pebble 5M
- Pebble 75M
- Pebble 1M (possibly)

We hope you're excited and enjoy the models!

Follow for more:
@Hoglet-33
basically-ai
  • 7 replies
Β·
Banaxi-TechΒ 
posted an update 2 days ago
view post
Post
2108
We have updated the BananaMind Base Bench leaderboard!
We now have these benchmark cards, they make it way easier to see which models are actually good!
We've also added the model advisor. It asks you what you want to use the model for and the parameter range and gives you the best model for your task!

Try it out at BananaMind/BananaMindBench-Leaderboard


And please give us a follow to BananaMind!
BananaMind

@Banaxi-Tech
  • 1 reply
Β·
etemizΒ 
posted an update 2 days ago
view post
Post
491
fine tuning going well, without breaking the model
  • 1 reply
Β·
GoktugDΒ 
posted an update 2 days ago
view post
Post
2291
πŸ‡ΉπŸ‡· We trained a 1B OCR model specifically for Turkish enterprise documents.

**Werea-DocOCR-1B v2**

The result surprised us:

LightOnOCR-2 base β†’ **64.2% CER**
Werea-DocOCR v1 β†’ **~8.1% CER**
Werea-DocOCR v2 β†’ **0.15% CER** πŸš€

Evaluated on a held-out 72-page test set across 12 Turkish document types and 3 different capture conditions.

πŸ“„ 12 Turkish enterprise document types
πŸ§ͺ 12,960 synthetic training pages
πŸ“± Digital + scanned + phone photos
πŸ“Š Tables β†’ structured Markdown
βš™οΈ Full-parameter fine-tuning
πŸ–₯️ Trained on a single RTX 3090

It handles:

β€’ e-Invoices
β€’ rental contracts
β€’ bank receipts
β€’ payroll documents
β€’ insurance policies
β€’ vehicle documents
β€’ official correspondence
β€’ trade registry documents
β€’ SGK-style tables
β€’ and more.

**Model πŸ€—**
Werea-co/Werea-DocOCR-1B

**Dataset πŸ“š**
Werea-co/werea-tr-doc-ocr-enterprise-v2

**Werea πŸ‡ΉπŸ‡·**
Werea-co


We're building open AI models from TΓΌrkiye.

This is just the beginning.

#HuggingFace #OCR #DocumentAI #TurkishAI #OpenSourceAI #ComputerVision
KlondikeDevΒ 
posted an update about 19 hours ago
view post
Post
595
The Small Language Model Consortium has reached 10 members!

Anybody is welcome to join, via two methods:

1. Starting a community discussion on the group’s README
2. Receiving an invite from me: @KlondikeDev

The Discord Community is also live! Feel free to join, even if you aren’t in the consortium, to discuss Small Language Models, or even just AI in general!

https://discord.gg/FngBKjja4

You’re more than welcome to apply!

β€” KlondikeDev