Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
celeba-sae-top_k-64-patches_only-layer_4-hook_resid_post-64-76 – AI Model by Prisma-Multimodal | AlphaNeural AI
You can deploy this model and start earning money today!
Prisma-Multimodal
/
celeba-sae-top_k-64-patches_only-layer_4-hook_resid_post-64-76
like
0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
CLIP Sparse Autoencoder Checkpoint
This model is a sparse autoencoder trained on CLIP's internal representations.
Model Details
Architecture
Layer
: 4
Layer Type
: hook_resid_post
Model
: open-clip:laion/CLIP-ViT-B-32-DataComp.XL-s13B-b90K
Dictionary Size
: 49152
Input Dimension
: 768
Expansion Factor
: 64
CLS Token Only
: False
Training
Training Images
: 324085
Learning Rate
: 0.0001
L1 Coefficient
: 0.0002
Batch Size
: 4096
Context Size
: 49
Performance Metrics
Sparsity
L0 (Active Features)
: 64.0000
Dead Features
: 0
Mean Log10 Feature Sparsity
: -3.1919
Features Below 1e-5
: 0.0000
Features Below 1e-6
: 0.0000
Mean Passes Since Fired
: 0.2603
Reconstruction
Explained Variance
: 0.7645
Explained Variance Std
: 0.0550
MSE Loss
: 0.0018
L1 Loss
: 0
Overall Loss
: 0.0018
Training Details
Training Duration
: 1175 seconds
Final Learning Rate
: 0.0000
Warm Up Steps
: 500
Gradient Clipping
: 1
Additional Information
Original Checkpoint Path
: /network/scratch/p/praneet.suresh/celeba_checkpoints/c3c29ea5-tinyclip_sae_16_hyperparam_sweep_lr/n_images_324169.pt
Wandb Run
:
https://wandb.ai/perceptual-alignment/celeba-sweep-topk-patches_all_layers/runs/ge5s3nf6
Random Seed
: 42