Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🤝
Open to Collab
358.3
TFLOPS
AbstractPhila
PRO
AbstractPhil
19
6
30
Follow
Alptekinege's profile picture
branikita's profile picture
marduk191's profile picture
104 followers
·
144 following
https://civitai.com/user/AbstractPhila
AbstractEyes
AI & ML interests
datasets, research papers, experimentation, vision, classification, text encoders, tokenization, llms, diffusion, distillation, and more.
Recent Activity
updated
a model
about 18 hours ago
AbstractPhil/alephllm-mini-beatrix-training
replied
to
their
post
about 20 hours ago
My apologies for the incorrect format for the AMOE arms from the experimental branch. They have been saving as torch objects. They are now correctly saving as safetensors format. My apologies for the inconvenience this may cause for you use. I will be modifying the codespaces to use the correct safetensors formats. After the first 20.9b tokens trained, the real experiments begins. Beatrix V3's first prepped-state modular command structure has been attached for dynamic training. These arms will exist as appendages for Beatrix - trained alongside with the trunk until the end of the run. These exist for experimental extraction, analysis, distillation experiments, memory experiments, mathematics experiments, and more. Each arm will be built along the chain for specific test cases. Expectation for each is already lined up and the outcomes are tested for, but the model still may face instability and must be monitored. As her first arm learns tinystories, she builds direct composite semantic structure throughout this system. Think of it like, the first higher-functioning cognition attachment. She's still very naïve and structurally unaware, so attaching new limbs is essentially extending a structure that is not yet finished forming. Nothing but fragments of issued information from an unknown source. In this case, this structure has been tested hundreds of times to ensure she will not simply collapse during training by having this attached.
replied
to
their
post
1 day ago
My apologies for the incorrect format for the AMOE arms from the experimental branch. They have been saving as torch objects. They are now correctly saving as safetensors format. My apologies for the inconvenience this may cause for you use. I will be modifying the codespaces to use the correct safetensors formats. After the first 20.9b tokens trained, the real experiments begins. Beatrix V3's first prepped-state modular command structure has been attached for dynamic training. These arms will exist as appendages for Beatrix - trained alongside with the trunk until the end of the run. These exist for experimental extraction, analysis, distillation experiments, memory experiments, mathematics experiments, and more. Each arm will be built along the chain for specific test cases. Expectation for each is already lined up and the outcomes are tested for, but the model still may face instability and must be monitored. As her first arm learns tinystories, she builds direct composite semantic structure throughout this system. Think of it like, the first higher-functioning cognition attachment. She's still very naïve and structurally unaware, so attaching new limbs is essentially extending a structure that is not yet finished forming. Nothing but fragments of issued information from an unknown source. In this case, this structure has been tested hundreds of times to ensure she will not simply collapse during training by having this attached.
View all activity
Organizations
AbstractPhil
's models
220
Sort: Recently updated
AbstractPhil/alephllm-mini-beatrix-training
Updated
about 1 hour ago
•
1
AbstractPhil/mega-liminal-lora
Text-to-Image
•
Updated
3 days ago
AbstractPhil/mini-beatrix-2s
Text Generation
•
0.2B
•
Updated
5 days ago
•
2.88k
AbstractPhil/mini-beatrix-1
Text Generation
•
0.1B
•
Updated
5 days ago
•
2.13k
•
1
AbstractPhil/mini-beatrix-2.5s
Text Generation
•
0.2B
•
Updated
6 days ago
•
249
AbstractPhil/mini-beatrix-3
Updated
7 days ago
AbstractPhil/geolip-bytelex
Updated
13 days ago
AbstractPhil/aleph-splat-0
Updated
23 days ago
AbstractPhil/rnn-cifar10-t0
Updated
about 1 month ago
AbstractPhil/alephlm-0
Feature Extraction
•
Updated
Aug 26
AbstractPhil/clip-vitb-mini-distilled
Image Feature Extraction
•
8.93M
•
Updated
Aug 24
•
70
AbstractPhil/alephlm-adopt-0
Text Generation
•
Updated
Aug 8
AbstractPhil/captionbert-8192-v2-b
Feature Extraction
•
58.3M
•
Updated
Aug 8
•
52
•
2
AbstractPhil/captionbert-8192-v2
Feature Extraction
•
58.3M
•
Updated
Aug 8
•
47
•
1
AbstractPhil/sd15-flow-lune
Text-to-Image
•
Updated
Aug 3
•
53
AbstractPhil/loss-manifest
Updated
Aug 2
AbstractPhil/geolip-bertenstein
Feature Extraction
•
Updated
Jul 31
AbstractPhil/geolip-vit-captionbank-coco
Image Feature Extraction
•
Updated
Jul 28
AbstractPhil/geolip-vit-base-x3
11.7M
•
Updated
Jul 28
•
10
AbstractPhil/geolip-vit-large-x3
78.3M
•
Updated
Jul 28
•
9
AbstractPhil/geolip-aleph-diffusion
Updated
Jul 25
•
2
AbstractPhil/geolip-aleph-qwen-3.5-0.8b-instruct
Updated
Jul 25
•
1
AbstractPhil/amoe-lora
Updated
Jul 21
AbstractPhil/aleph-diffusion-adapters
Updated
Jul 19
AbstractPhil/qwen3.5-0.8b-relay-caption
Updated
Jul 18
AbstractPhil/geolip-aleph-qwen
Updated
Jul 15
AbstractPhil/geolip-aleph-differentiation
Updated
Jul 12
AbstractPhil/anima-90k
Updated
Jul 7
•
1
AbstractPhil/geolip-aleph-lm
Text Generation
•
Updated
Jul 1
•
1
AbstractPhil/qwen-benchmark
Updated
Jun 30
Previous
1
2
3
...
8
Next