GitHub · by RJMSWD
QwenJev
A Jev-style conversion of Qwen-3.5-4B, at 0.169 s per frame.
The Qwen-3.5-4B vision-language model converted to work like Jev, reaching 0.169 seconds per frame for recognition. The author notes the same approach transfers to the other Qwen models.
Open the source