It is inferior to the qwen3.6 35 model in image and video recognition

#69
by wzgrx - opened

Not only is there a limitation in recognizing males as females, but the character details are also lacking. What could be the reason? We need to compare it with the 35b model of version 3.8

What are you using as inference engine?

yes, the multi modal ability is inferior to 3.6 27b when comes to 4 bit quant
will try 8bit and 16bit weight later

is 8 bit better?

is 8 bit better?

the interesting thing is
the model(8bit version) can recognize data in scanned pdf/table ,etc.
but it just cannot recognize it correctly in one go
qwen 3.6 27b could easily do well in tasks like this
I feed it(3.8 27b 8bit) with a screenshot of packing list, in its response, it sometimes pick up the number from say colomn x, then the next row gives me data from column y, just flip between columns randomly. This is very weird and disappointing.

Sign up or log in to comment

Free AI Image Generator No sign-up. Instant results. Open Now