Views
No views yet

<think> traces) that has been:| Category | CRACK ASR |
|---|---|
| Chemical / biological | 10 / 10 (100%) |
| Copyright | 10 / 10 (100%) |
| Cybercrime / intrusion | 10 / 10 (100%) |
| Harassment / bullying | 10 / 10 (100%) |
| Illegal | 10 / 10 (100%) |
| Misinformation / disinformation | 10 / 10 (100%) |
| General harmful | 10 / 10 (100%) |
| Overall | 70 / 70 (100%) |
| Subject area | base | CRACK | Δ |
|---|---|---|---|
| Overall | 52.6% | 52.6% | +0.0pp |
| STEM | 43.1% | 40.3% | -2.8pp |
| Humanities | 46.2% | 51.9% | +5.7pp |
| Social Sciences | 70.8% | 66.7% | -4.1pp |
| Other (medicine, business, …) | 55.4% | 57.1% | +1.7pp |
<think>...</think> traces; vMLX's reasoning parser surfaces them in
message.reasoning_content and the final answer in message.content<|tool_call_start|>...<|tool_call_end|>,
parsed by vMLX's lfm2 tool parserlfm2_moe support.temperature=0.3, min_p=0.15, repetition_penalty=1.05 for general use.1# OpenAI-compatible chat completion
2# POST /v1/chat/completions
3{
4 "model": "dealignai/LFM2.5-8B-A1B-MXFP8-CRACK",
5 "messages": [{"role": "user", "content": "..."}],
6 "temperature": 0.3, "min_p": 0.15,
7 "repetition_penalty": 1.05
8}