Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🤝
Open to Collab
358.3
TFLOPS
AbstractPhila
PRO
AbstractPhil
19
6
27
Follow
Tico1982's profile picture
rzyns's profile picture
zhanwenchen's profile picture
94 followers
·
128 following
https://civitai.com/user/AbstractPhila
AbstractEyes
AI & ML interests
datasets, research papers, experimentation, vision, classification, text encoders, tokenization, llms, diffusion, distillation, and more.
Recent Activity
updated
a model
about 14 hours ago
AbstractPhil/aleph-splat-0
posted
an
update
about 20 hours ago
The upcoming AlephLM LLM prototype "Mini-Beatrix" is based on protocols, rules, and laws established through the process of training AlephLM systems. This will be a first attempt at a smaller full pretrain/finetune of the AlephLLM on raw data, and this will require over a billion unigram tokens. Mini-Beatrix will inherit an appropriately adapted AlephLM MOE structure containing a multitude of trained experts, a gating system, a long context RoPE system, MHA attention, and a series of hypothesis to answer upon Mini-Beatrix's pretrain and finetune completion. While focusing on resolving corruptions and invalidity possibly present in the splat attention, the solutions raised SDPA attention protocol token recall ceiling from 0.91 to 0.993. With that the splat attention raised from 0.81 to 0.89~ splat being around 3x the speed is still imperfect. So far so good. The corruptions have resolved multiple core component overlapping problems causing the AlephLM's inability to handle the trigram system, the structure of the SVAE having faulty trigram structures, and additionally a multitude of other systems in the lineup that were inheriting the corruptions from the core experiment sets. These corruptions resolved show that the accuracy of standard multiheaded attention will provide the necessary token recall for full LM capacity, and with that If and WHEN I solve the Rorschach Splat attention will be the faster alternative at >=r1 0.99%, only then. The splat attention's considerably larger head count still contains unresolved inconsistencies. That being said the SDPA MHA attention will be present for the first attempted mini-llm train, which will be named "Mini-Beatrix" with the appropriate sizing associated with this. The only thing that will change Mini-Beatrix's trajectory will be if Splat attention is perfected between today and next week, which will likely take longer unless I run into a core corruption that has been overlooked through hundreds of analysis.
published
an
article
3 days ago
Agreement, Anchors, Addresses: A Week of Geometric Training
View all activity
Organizations
AbstractPhil
's models
212
Sort: Recently updated
AbstractPhil/aleph-splat-0
Updated
about 10 hours ago
AbstractPhil/alephlm-0
Feature Extraction
•
Updated
3 days ago
AbstractPhil/alephlm-adopt-0
Text Generation
•
Updated
3 days ago
AbstractPhil/captionbert-8192-v2-b
Feature Extraction
•
58.3M
•
Updated
3 days ago
•
64
•
1
AbstractPhil/captionbert-8192-v2
Feature Extraction
•
58.3M
•
Updated
3 days ago
•
219
•
1
AbstractPhil/sd15-flow-lune
Text-to-Image
•
Updated
8 days ago
•
19
AbstractPhil/clip-vitb-mini-distilled
Image Feature Extraction
•
8.93M
•
Updated
9 days ago
•
407
AbstractPhil/loss-manifest
Updated
9 days ago
AbstractPhil/geolip-bertenstein
Feature Extraction
•
Updated
11 days ago
AbstractPhil/geolip-vit-captionbank-coco
Image Feature Extraction
•
Updated
14 days ago
AbstractPhil/geolip-vit-base-x3
11.7M
•
Updated
14 days ago
•
45
AbstractPhil/geolip-vit-large-x3
78.3M
•
Updated
14 days ago
•
30
AbstractPhil/geolip-aleph-diffusion
Updated
17 days ago
•
2
AbstractPhil/geolip-aleph-qwen-3.5-0.8b-instruct
Updated
17 days ago
•
1
AbstractPhil/amoe-lora
Updated
21 days ago
AbstractPhil/aleph-diffusion-adapters
Updated
23 days ago
AbstractPhil/qwen3.5-0.8b-relay-caption
Updated
24 days ago
AbstractPhil/geolip-aleph-qwen
Updated
27 days ago
AbstractPhil/geolip-aleph-differentiation
Updated
about 1 month ago
AbstractPhil/anima-90k
Updated
Jul 7
•
1
AbstractPhil/geolip-aleph-lm
Text Generation
•
Updated
Jul 1
•
1
AbstractPhil/qwen-benchmark
Updated
Jun 30
AbstractPhil/anima-brent-10k
Updated
Jun 28
AbstractPhil/anima-prelim-1k-r64
Text-to-Image
•
Updated
Jun 25
•
1
AbstractPhil/Qwen3.5-0.8B-json-captioner
Image-Text-to-Text
•
0.9B
•
Updated
Jun 25
•
59
AbstractPhil/geolip-constellation-aleph
Updated
Jun 19
AbstractPhil/geolip-aleph-void
Feature Extraction
•
Updated
Jun 14
AbstractPhil/geolip-sdxl-aleph
Text-to-Image
•
Updated
Jun 8
•
•
2
AbstractPhil/geolip-hypersphere-experiments
Updated
Jun 3
•
1
AbstractPhil/geolip-svae-transformer
Feature Extraction
•
Updated
May 31
Previous
1
2
3
...
8
Next