Views
No views yet
AWQModifier applied to all Linear layers of the language model. The vision tower and lm_head are kept in BF16.HuggingFaceH4/ultrachat_200k, max sequence length 2048W4A16_ASYMlm_head, visual.* (vision tower)vllm serve <this-repo> --trust-remote-code=falseExtract all readable content from the image in natural human reading order and output the result as a single Markdown document. Format formulas as LaTeX. Format tables as HTML: <table>...</table>. Preserve the original text without translation.