Vladimir Bataev، Lilit Grigoryan، Andrei Andrusenko، Nikolay Karpov و 2 نفر دیگر
زمینهسازی برای سیستمهای تشخیص گفتار خودکار (ASR) در محیطهای عملیاتی، اهمیتی حیاتی دارد؛ بهگونهای که جملاتِ ارائهشده توسط کاربر باید تحت محدودیتهای زمانی سخت، با دقت بالا تشخیص داده شوند. اگرچه بسیاری از روشهای تأ…
Shiao Xie، Siyu Chen، Jianwei Lv، Bo Yuan و 2 نفر دیگر
Personalized interpretation of medical reports has emerged as an increasingly important need among patients. Addressing this need requires both evidence-grounded medical factuality and context-dependent patient communica…
Narges Ahmadi، Yubo Jiao، Jônatas Augusto Manzolli، Jiangbo Yu و 1 نفر دیگر
تحقیقات در زمینه رفتار مسافران، دادههای دیجیتال را به طور متمرکزتری با مدلسازی پیشبینی ترکیب میکنند، اما این مراحل معمولاً به طور جداگانه توسعه و ارزیابی میشوند. این مطالعه یک فرآیند سهمرحلهای را پیشنهاد میکند که…
Yucheng Jiang، Zora Zhiruo Wang، Ruishi Chen، Diyi Yang
Naturalistic computer-use traces, passively recorded screenshots and mouse or keyboard actions, are a valuable resource for deriving symbolic, auditable, and reusable models of how everyday work is done. Such models matt…
Yizhe Chi، Wenyi Li، Deyao Hong، Xiaoqiu Wang و 6 نفر دیگر
Recursive self-improvement (RSI) asks whether an AI system can improve the process that produces AI systems, so that the next system inherits the improvement. That process is the training algorithm: a better objective or…
Adam Fisch، Shubhendu Trivedi، Fantine Huot، William W. Cohen و 4 نفر دیگر
Heterogeneous AI systems composed of multiple models, architectures, harnesses, or inference-time settings can improve quality and efficiency by routing queries to the specialist who can answer most effectively at the lo…
Fengqing Jiang، Yite Wang، Boyi Liu، Zhaoyang Wang و 4 نفر دیگر
Mid-training is increasingly recognized as a critical stage for shaping the capabilities of large language models. Recent work has shown that targeted mid-training can strengthen reasoning-intensive abilities such as mat…
آیا یک مدل زبانی خود را بهبود داده است یا خیر؟ به طور تدریجی، این مدل بر اساس دقت متوسط عمل نمیکند؛ بلکه در مورد نحوه مواجهه با مسائل فردی و نقاط ضعف آن مشکل دارد. ردیابی این تغییرات به معنای تفریق دو برآورد نویزآمیز اس…
1. بهبود جریان و خوانایی
2. ساختار جملهای طبیعیتر در فارسی
3. ثبات در اصطلاحات
4. تصحیح خطاهای ترجمه
مقررات:
- حفظ اصطلاحات و فرمولهای فنی
- حفظ ساختار بخشها
- حفظ استنادها و منابع
فقط متن فارسی بهبود یافته را ارائ…
Large language model (LLM) agents can induce skills from completed tasks and reuse them later to grow more capable with experience. In practice, induced skills may transfer unreliably and can even harm the agent that ret…