vision language model skills

9 tagged vision language model, measured the same way as everything else here.

Browse within: llava 8llava-onevision 8mllm 8qwen3 8

merge-ov2

02

EvolvingLMMs-Lab/LLaVA-OneVision-2

Skill Claude CodeCodex

Bilingual guide for merging ViT + LLM into LlavaOnevision2 HF checkpoint and validating weight/inference consistency.

1.2k 3d ago A 29 tokens original Apache-2.0

EvolvingLMMs-Lab/LLaVA-OneVision-2

Skill Claude CodeCodex

Bilingual guide for the OFFLINEPACKINGBMR and OFFLINEPACKEDDATA environment variables that control LLaVA-OneVision2 training-side packing — what each gate does, why both must be enabled together, MBS=1 requirement, and the dead OFFLINEPACKINGVQA branch.

1.2k 3d ago A 65 tokens original Apache-2.0

screenclaw

04

GinSing1226/ScreenClaw

Skill Claude CodeCodex

Desktop software automation through screenshots and coordinate grids. It can repeat recorded actions as reusable templates without needing the software to provide an API or command line.

48 2mo ago A 298 tokens original MIT