طراحی سامانه ارزیابی مقاوم مبتنی بر LLM برای هوش مصنوعی عامل در کشف دارو با تراز انسانی
مدلهای زبانی AI 73/100 تحلیل هوشمند
خلاصه خبر واقعیت از منبع
arXiv:2608.21057v1 Announce Type: new Abstract: Agentic large language model (LLM) systems are reshaping scientific workflows in chemistry and drug discovery, but evaluating their open-ended, tool-augmented outputs remains a fundamental bottleneck. Reference-based metrics such as BLEU and ROUGE fail to capture semantic correctness, while expert human evaluation does not scale to the iteration speed these systems demand. The LLM-as-a-Judge paradigm has emerged as a scalable a…
تحلیل راهبردی تحلیل / هوشمندی
سیستمهای هوش مصنوعی عامل مبتنی بر مدلهای زبان بزرگ در کشف دارو تحول ایجاد کردهاند، اما ارزیابی خروجیهای پیچیده آنها چالشی اساسی است.
چرا مهم است؟
—
اثر احتمالی برای ایران
متوسط
- کلی: 63
- فناوری: 100
- اقتصادی: 60
- صنعتی: 45
- توانمندی ملی: 23
حوزه راهبردی
مدلهای زبانی
موجودیتهای مرتبط
عامل هوشمند · technology هوش عاملمحور · technology مدل زبانی بزرگ · technology تراشه AI · technology