Nanami_Chiaki (@wjw030515) 在 做了个大模型视觉能力检测的小工具 中发帖
SkalskiP on X: "I’m both impressed and disappointed by Gemini 3.6 Flash faster, cheaper, and uses fewer tokens then Gemini 3.5 Flash. noticeably worse at object detection. often returns one general box instead of several precise ones. feels lazy. prompt: detect banana tree ↓ deep dive https://t.co/YREla5HVHx" / X 看到有人分享这条帖子,于是让ai搓了一个单html工具,能够让ai给画面中的物体打标并渲染出来。
index.7z (6.6 KB)
简单测试了几个模型,结果如下:
...