Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
TinyLoRA-TexasHoldEm-Llama-3.2-1B-Instruct – AI Model by neopolita | AlphaNeural AI
You can deploy this model and start earning money today!
neopolita
/
TinyLoRA-TexasHoldEm-Llama-3.2-1B-Instruct
like
0
poker
texas-holdem
fine-tuned
lora
2602.04118
meta-llama/Llama-3.2-1B-Instruct
adapter
llama3.2
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
TinyLoRA-TexasHoldEm-Llama-3.2-1B-Instruct
Fine-tuned Llama 3.2 1B Instruct model for Texas Hold'em poker decisions with
TinyLoRA
. The adapter size is 470KB!
Training
Base model:
meta-llama/Llama-3.2-1B-Instruct
Dataset:
RZ412/PokerBench
Method:
TinyLoRA fine-tuning with unsloth-mlx
LoRA config:
r=256 (svd), u=1024, target_modules=[q_proj, k_proj, v_proj, o_proj, gate_proj, up_proj, down_proj]
Training data:
50k preflop + 50k postflop samples
Performance
Evaluated on PokerBench test sets:
Preflop
Postflop
Base Model
7%
13%
Fine-tuned
36%
60%