AI & ML interests

Models and API early access previews for KoFi donators

Recent Activity

appvoidΒ  updated a Space 4 days ago
CEAMFA/README
appvoidΒ  updated a Space 9 days ago
CEAMFA/palmer
appvoidΒ  published a Space 9 days ago
CEAMFA/palmer
View all activity

appvoidΒ 
posted an update about 22 hours ago
view post
Post
547
Random corporate secret of tonight:

Try overfitting a tiny model on billions of high-quality datapoints: you can't. You can do 100 epochs and see the model still improving.

You're welcome.
  • 1 reply
Β·
appvoidΒ 
posted an update 2 days ago
view post
Post
711
- GLM 5.2
- Flux 3
- New Qwen model
- New small model leaderboards
- Lots of people finetuning smol models.
- Some even under 12 year olds clauders are here (was not on my bingo card this year)
- ChatGPT's Sol became a lot faster this week
- LFM2.5 2.6b
- Kimi K3 (though only a few will run it)
- New Ling 3.0 Tiny
- New video model that is making south park videos?
- Deepseek v4 flash being more honest than bigger models
- The new model from meta


Everything Everywhere All At Once
  • 1 reply
Β·
appvoidΒ 
updated a Space 4 days ago
appvoidΒ 
posted an update 4 days ago
view post
Post
2358
I don't know if it was us or one of you guys or maybe all of us at once but lately we have seen a finetuning/pretraining explosion of models below 200m params and we can't be more happy about it keep coming tinkerers all of this is possible because of you!
  • 8 replies
Β·
appvoidΒ 
posted an update 5 days ago
view post
Post
2883
If you want small models to be great again you should give a follow to people like @Banaxi-Tech or @Datdanboi25

These guys are rocking it with small models lately.

(They are not paying me to say that)
  • 22 replies
Β·
appvoidΒ 
posted an update 7 days ago
view post
Post
176
byte-level
deep layers
diverse data
compact size
overfitting
grokking

I know the next best small model is somewhere in the intersection of these features.
  • 1 reply
Β·
appvoidΒ 
posted an update 8 days ago
view post
Post
891
Do you prefer AI companies/individuals to release half-baked models weekly or an overpowered model a month?
  • 2 replies
Β·
appvoidΒ 
published a Space 9 days ago
appvoidΒ 
posted an update 9 days ago
view post
Post
1350
Giving free early access to the gguf for some of you today! Tell me what you think.

CEAMFA/palmer-007-preview
appvoidΒ 
posted an update 14 days ago
view post
Post
614
i love reinforcement learning
  • 1 reply
Β·
appvoidΒ 
published a Space 16 days ago
appvoidΒ 
posted an update 16 days ago
view post
Post
1411
A Small Model is All You Need. Meet palmer-006 (90M)

After 3 years of experiments, we are finally releasing our flagship tiny model: **palmer-006**.

If you are building for edge hardware, SBCs (Raspberry Pi, etc.), or low-power devices, this is for you. Inspired by Andrej Karpathy's idea of a self-contained "cognitive core," we wanted to see how much power we could pack into a sub-100M parameter footprint.

🧠 **How we "Palmerized" it:**
We believe in starting our experiments with the absolute strongest baseline possible.
1. Light fine-tuning on highly curated data
2. Model merging
3. Another light fine-tuning round
4. Adjusted Mamba for maximum token speed ⚑️

⚠️ *Note: This is a foundational language model. It has not been instruction-tuned yet!*

Also, since this needs instruction tuning next to become a chat assistantβ€”**what dataset would you recommend we use for the instruct tune?**

---
πŸ”— **Quick Links & Info:**

* **License:** Open for research, education, hobby, and modification! (For commercial use/hosted APIs, shoot an email to nosoyhackercodigo@gmail.com. *PS: Donators can claim a free commercial license!*)

* **Attribution:** Built using AI tech from the Technology Innovation Institute (TII).

Can't wait to see what you build at the edge. Let me know your prompt completions below! πŸ‘‡

appvoid/palmer-006
  • 10 replies
Β·
appvoidΒ 
posted an update 17 days ago
view post
Post
204
...for tomorrow.
appvoidΒ 
posted an update 21 days ago
view post
Post
3338
Two big projects are open sourced soon. Get ready...
  • 11 replies
Β·
appvoidΒ 
posted an update 24 days ago
view post
Post
223
If you make cool smol πŸ€–πŸ€ models (below 0.5b parameters), leave a reply and I will follow you! I'm serious, you don't need to follow me at all just share something through the replies and I (and potentially more people) will follow you (if your models are decent ofc).
  • 11 replies
Β·
appvoidΒ 
posted an update 25 days ago
view post
Post
4661
if you are a tinkerer of small language models and want to stay ahead of what small models can do, follow me!!! seriously, start following people that actually still makes small models

i've made one recently btw

also, i'm keeping an eye on AxiomicLabs leaderboard, looks like the only current alternative to check where the things are going to

though, between us, i think they should add agentic/tool use benchmarks there

anyways,


enjoy!

appvoid/a-cool-model
  • 19 replies
Β·
appvoidΒ 
posted an update 26 days ago
view post
Post
114
A huge amount of large synthetic datasets on huggingface looks surprisingly like templates, that might be one of the main reasons open models might not be as good as other models, we need more people to create smaller, human-curated datasets instead of lazily sending millions of requests to large models for us to fulfill.
  • 4 replies
Β·