๐ค quecto mode
appvoid
appvoid
AI & ML interests
singularity through byte-level tokens and small language models - creating applications and ideas out of the void
Recent Activity
published a dataset about 9 hours ago
appvoid/void-byte-data updated a dataset about 12 hours ago
appvoid/void-byte-data liked a Space about 13 hours ago
BananaMind/BananaMindBench-LeaderboardOrganizations
posted an update about 23 hours ago
posted an update 12 days ago
Post
122
LLMS for edge devices?
What if we go the reliable route instead of the speed route?
What if we make it run on sbcs with few megabytes available?
That's the idea for the next model.
Keep in tune.
What if we go the reliable route instead of the speed route?
What if we make it run on sbcs with few megabytes available?
That's the idea for the next model.
Keep in tune.
congrats!
posted an update 15 days ago
Post
137
We trained a 10.9M byte-level recurrent Transformer on L3 and L6. (Loop 3 and Loop 6)
Yet L4/L5 improved too, L8 held up, and the L3โL6 gain grew during training.
Same weights. More compute. Better predictions.
This is a new architecture for effective compute after several steps beyond original training!
We mixed and matched components like time and mhc into an ouro-like byte-level language model and the result is BET, a byte-level step-elastic transformer that can run computation steps without significant degradation.
One of the coolest parts of this training was discovering how Gradient Descent decided to use the first layer as what we would consider a scratchpad! Totally destroyed for the decoder but somehow makes total sense for the next layer!
I believe looped-transformers are the future of edge computing and this is a first step towards it.
Blogpost: https://medium.com/@appvoidofficial/byte-level-elasticity-182fe2ed1d2f
appvoid/bet-10m
Yet L4/L5 improved too, L8 held up, and the L3โL6 gain grew during training.
Same weights. More compute. Better predictions.
This is a new architecture for effective compute after several steps beyond original training!
We mixed and matched components like time and mhc into an ouro-like byte-level language model and the result is BET, a byte-level step-elastic transformer that can run computation steps without significant degradation.
One of the coolest parts of this training was discovering how Gradient Descent decided to use the first layer as what we would consider a scratchpad! Totally destroyed for the decoder but somehow makes total sense for the next layer!
I believe looped-transformers are the future of edge computing and this is a first step towards it.
Blogpost: https://medium.com/@appvoidofficial/byte-level-elasticity-182fe2ed1d2f
appvoid/bet-10m
reacted to KlondikeDev's post with โค๏ธ 17 days ago
Post
2151
The SLM Consortium has begun work on a safety dataset for Small Language Models, with the creation of the dataset being headed by @wayneworkman2012
The dataset will focus on refusals and redirects surrounding dangerous or extreme sexual content, designed to be shaped sized appropriately for SLMs, without significantly lowering benchmark performance.
More info will be out soon!
slmconsortium
The dataset will focus on refusals and redirects surrounding dangerous or extreme sexual content, designed to be shaped sized appropriately for SLMs, without significantly lowering benchmark performance.
More info will be out soon!
replied to their post 19 days ago
Well I donโt care if itโs AGI or not I just find it super helpful in real practice
It would be for me if my credits wouldn't drain even light mode ๐ฅฒ
replied to their post 19 days ago
it is not AGI
For me it is not, but it's pretty close for me though
replied to their post 20 days ago
Ill wait for the glm distilled xD
I think we are years away though
replied to their post 20 days ago
Wild if true
replied to their post 22 days ago
Feeling the same :)
replied to Banaxi-Tech's post 26 days ago
Imagine having some kind of server so it always feels updated with your latest integrated models!
reacted to Banaxi-Tech's post with ๐ 26 days ago
Post
3452
Introducing BananaMindOS 3.0
- Complete modern UI redesign
- Adds support for Qwen3.5 0.8B, LFM2.5 230M,350M, SmolLM2 360M, Gemma 3 270M.
- Adds Q7,Q6,Q5,Q3,Q1 quantization formats with a easy to use precision slider
- And more!
The new UI includes:
- New 1024ร768 High Quality interface.
- Photographic QOI background.
- Transparent BananaMind, CPU, cube, mouse, and Send icons.
- Proper bitmap cursor.
- Rounded translucent panels and cards.
- Modern model-loading progress window.
- Redesigned inference screen with response and prompt panels.
- Localized redraws for the cursor, clicks, loading progress, and precision slider.
Notice: Qwen3.5 0.8B currently generates garbled text, it will be fixed tomorrow.
See it for yourself
Now Available at https://github.com/BananaMind/BananaMindOS
Prebuild ISOs coming soon!
(also press ? + G if you want to load try to load a 6MB RAM model on 5MB may break)
- Complete modern UI redesign
- Adds support for Qwen3.5 0.8B, LFM2.5 230M,350M, SmolLM2 360M, Gemma 3 270M.
- Adds Q7,Q6,Q5,Q3,Q1 quantization formats with a easy to use precision slider
- And more!
The new UI includes:
- New 1024ร768 High Quality interface.
- Photographic QOI background.
- Transparent BananaMind, CPU, cube, mouse, and Send icons.
- Proper bitmap cursor.
- Rounded translucent panels and cards.
- Modern model-loading progress window.
- Redesigned inference screen with response and prompt panels.
- Localized redraws for the cursor, clicks, loading progress, and precision slider.
Notice: Qwen3.5 0.8B currently generates garbled text, it will be fixed tomorrow.
See it for yourself
Now Available at https://github.com/BananaMind/BananaMindOS
Prebuild ISOs coming soon!
(also press ? + G if you want to load try to load a 6MB RAM model on 5MB may break)
replied to their post 26 days ago
byte-level tokenizer is the part that caught my attention, keep it up sir!
replied to their post 26 days ago
Let's go!!!
replied to their post 26 days ago
Yes sir!
replied to their post 26 days ago
Do you mean like the architecture? Could you point at the model?
replied to their post 26 days ago
Waiting for your new models sir
replied to their post 26 days ago
Hey @appvoid
Are you gonna press it?
Done. hahaha bots are becoming more of a thing here lately.
