Bigger AI does not mean less biased AI, a test of 194 models suggestsبررسی 194 مدل نشان میدهد بزرگتر شدن AI لزوماً سوگیری کمتری به همراه ندارد
Making picture-and-text AI models bigger did not stop them relying on misleading clues. Across 194 public models, the bigger versions gained 2.6% on a standard picture test but lost 4.2% on the hardest fairness test.بزرگتر کردن مدلهای AI که تصویر و متن را با هم پردازش میکنند، مانع از تکیهشان بر سرنخهای گمراهکننده نشد. در میان 194 مدل عمومی، نسخههای بزرگتر در یک آزمون تصویری استاندارد 2.6% پیشرفت داشتند، اما در سختترین آزمون انصاف 4.2% افت کردند.
ترجمهٔ ماشینی است؛ برای دقت به متن اصلی انگلیسی مراجعه کنید.
Why it matters
These models sit behind image search, photo filtering on social media, and tools that make pictures from words. A model that secretly judges a car by its background, or hair colour by a person's gender, will fail exactly on the people and things that do not fit the usual pattern. Right now, companies mostly pick models by one overall accuracy score, and this work suggests that score says little about such failures. The researchers found that the size and care put into the training pictures mattered far more than raw model size, though they only checked two kinds of bias, so the pattern may not hold everywhere.این مدلها پشت جستوجوی تصویر، فیلترهای عکس در شبکههای اجتماعی، و ابزارهایی که از روی متن تصویر میسازند قرار دارند. مدلی که پنهانی یک خودرو را بر اساس پسزمینهاش قضاوت کند، یا رنگ مو را بر اساس جنسیت فرد تشخیص دهد، دقیقاً همانجایی شکست میخورد که افراد و اشیا با الگوی معمول همخوانی ندارند. در حال حاضر شرکتها عمدتاً مدلها را بر اساس یک نمرهی کلی accuracy انتخاب میکنند، و این پژوهش نشان میدهد این نمره چیز زیادی دربارهی چنین شکستهایی نمیگوید. پژوهشگران دریافتند که اندازه و دقتِ کار روی تصاویر آموزشی، بسیار بیشتر از اندازهی خام مدل اهمیت داشت، هرچند آنها فقط دو نوع سوگیری را بررسی کردند، پس ممکن است این الگو همهجا صدق نکند.
Who's behind it: Ioannis Sarridis, Ioannis Kompatsiaris and Symeon Papadopoulos, Centre for Research and Technology Hellas (CERTH), Greece. Funded by the European Union's Horizon Europe research programme.
Summary by the Lemma AI · how we grade