Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🤝
Open to Collab
358.3
TFLOPS
AbstractPhila
PRO
AbstractPhil
19
6
27
Follow
joonsoo-me's profile picture
CosmicCrafter's profile picture
armant1372's profile picture
94 followers
·
127 following
https://civitai.com/user/AbstractPhila
AbstractEyes
AI & ML interests
datasets, research papers, experimentation, vision, classification, text encoders, tokenization, llms, diffusion, distillation, and more.
Recent Activity
posted
an
update
about 1 hour ago
The upcoming AlephLM LLM prototype "Mini-Beatrix" is based on protocols, rules, and laws established through the process of training AlephLM systems. This will be a first attempt at a smaller full pretrain/finetune of the AlephLLM on raw data, and this will require over a billion unigram tokens. Mini-Beatrix will inherit an appropriately adapted AlephLM MOE structure containing a multitude of trained experts, a gating system, a long context RoPE system, MHA attention, and a series of hypothesis to answer upon Mini-Beatrix's pretrain and finetune completion. While focusing on resolving corruptions and invalidity possibly present in the splat attention, the solutions raised SDPA attention protocol token recall ceiling from 0.91 to 0.993. With that the splat attention raised from 0.81 to 0.89~ splat being around 3x the speed is still imperfect. So far so good. The corruptions have resolved multiple core component overlapping problems causing the AlephLM's inability to handle the trigram system, the structure of the SVAE having faulty trigram structures, and additionally a multitude of other systems in the lineup that were inheriting the corruptions from the core experiment sets. These corruptions resolved show that the accuracy of standard multiheaded attention will provide the necessary token recall for full LM capacity, and with that If and WHEN I solve the Rorschach Splat attention will be the faster alternative at >=r1 0.99%, only then. The splat attention's considerably larger head count still contains unresolved inconsistencies. That being said the SDPA MHA attention will be present for the first attempted mini-llm train, which will be named "Mini-Beatrix" with the appropriate sizing associated with this. The only thing that will change Mini-Beatrix's trajectory will be if Splat attention is perfected between today and next week, which will likely take longer unless I run into a core corruption that has been overlooked through hundreds of analysis.
updated
a model
about 19 hours ago
AbstractPhil/aleph-splat-0
published
an
article
2 days ago
Agreement, Anchors, Addresses: A Week of Geometric Training
View all activity
Organizations
AbstractPhil
's datasets
82
Sort: Recently updated
AbstractPhil/captionbert-8192-v2-consensus
Updated
9 days ago
•
201
AbstractPhil/conceptual-captions-12m-webdataset-berts
Viewer
•
Updated
9 days ago
•
32.3M
•
532
•
1
AbstractPhil/bulk-cc12m-features
Viewer
•
Updated
10 days ago
•
121M
•
3.86k
AbstractPhil/tower-probes-results
Viewer
•
Updated
20 days ago
•
17
•
140
AbstractPhil/qwen-deepfashion-fused
Viewer
•
Updated
29 days ago
•
122k
•
2.27k
•
1
AbstractPhil/qwen-synth-characters-fused
Viewer
•
Updated
Jul 10
•
42.7k
•
1.73k
AbstractPhil/qwen-synth-characters-100-json-test
Viewer
•
Updated
Jul 10
•
1k
•
63
AbstractPhil/anima-brent-90k-cache
Updated
Jul 5
•
46
AbstractPhil/qwen-synth-characters
Viewer
•
Updated
Jul 3
•
61k
•
122
AbstractPhil/qwen-deepfashion
Viewer
•
Updated
Jul 3
•
160k
•
355
AbstractPhil/diffusion-pipe-cache-test1
Viewer
•
Updated
Jun 27
•
8.92k
•
65
AbstractPhil/anima-90k-cache
Updated
Jun 26
•
107
AbstractPhil/diffusion-pretrain-set-ft1
Viewer
•
Updated
Jun 23
•
1.46M
•
1.58k
•
1
AbstractPhil/diffusion-pretrain-set-ft1-1024
Viewer
•
Updated
Jun 11
•
1.14M
•
716
AbstractPhil/sdxl-qwen-phase1-cache
Viewer
•
Updated
Jun 6
•
86k
•
310
AbstractPhil/geolip-sdxl-fid-scoring
Viewer
•
Updated
Jun 5
•
2.8k
•
100
AbstractPhil/sdxl-qwen-phase0
Viewer
•
Updated
Jun 4
•
86k
•
240
•
3
AbstractPhil/IMDB-PUBLIC-SCRAPED
Preview
•
Updated
May 19
•
64
•
1
AbstractPhil/ldhnam-deepfashion_controlnet
Viewer
•
Updated
May 19
•
26k
•
24
AbstractPhil/ffhq_flux_latents_repaired
Viewer
•
Updated
May 19
•
40.8k
•
221
AbstractPhil/synthetic-characters
Viewer
•
Updated
May 19
•
149k
•
323
AbstractPhil/CN_pose3D_V10_512
Viewer
•
Updated
May 19
•
66.5k
•
69
AbstractPhil/CN_pose3D_V7_512
Viewer
•
Updated
May 19
•
255k
•
483
AbstractPhil/synthetic-object-relations-json
Viewer
•
Updated
May 18
•
5k
•
43
AbstractPhil/cc-task1-json
Preview
•
Updated
May 18
•
42
AbstractPhil/cc-prompts-sharded
Viewer
•
Updated
May 15
•
3.32M
•
10
AbstractPhil/json-coco-format
Viewer
•
Updated
May 14
•
129k
•
198
AbstractPhil/svae-freckles-4096-cifar10
Viewer
•
Updated
Apr 10
•
60k
•
57
AbstractPhil/ryan-spearman-prepared-features
Viewer
•
Updated
Mar 27
•
1
•
64
AbstractPhil/bertenstein-v1
Viewer
•
Updated
Mar 7
•
37.4k
•
1.24k
Previous
1
2
3
Next