Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
worldpolicy-grpo-3b-merged – AI Model by krishpotanwar | AlphaNeural AI
You can deploy this model and start earning money today!
krishpotanwar
/
worldpolicy-grpo-3b-merged
like
0
transformers
safetensors
llama
text-generation
merged
worldpolicy
grpo
conversational
unsloth/Llama-3.2-3B-Instruct
finetune
text-generation-inference
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
WorldPolicy GRPO 3B Merged
Standalone merged model for WorldPolicy debate inference.
Base:
unsloth/Llama-3.2-3B-Instruct
Adapter:
krishpotanwar/worldpolicy-grpo-3b
Context capped for serving:
4096
tokens