0

دانلود کتاب هنر پست آموزش (Post-Training)، راهنمای عملی برای مهندسان و توسعه‌دهندگان هوش مصنوعی

  • عنوان کتاب: The Craft of Post-Training -A Practical Guide for AI Engineers and Developers
  • نویسنده: Von Csefalvay, Chris
  • حوزه: توسعه هوش مصنوعی
  • سال انتشار: 2026
  • تعداد صفحه: 434
  • زبان اصلی: انگلیسی
  • نوع فایل: pdf
  • حجم فایل: 17.1 مگابایت

خاخام یهودا لو بن بزالل، معروف به مهارال پراگ، یک محقق، فیلسوف و تا حدودی عارف یهودی قرن شانزدهم بود. طبق افسانه‌ها، او با مشاهده آزار و اذیت قوم خود، از گل رس جمع‌آوری‌شده در سواحل رودخانه ولتاوا، محافظی برای آنها ساخت. او گل را به شکل یک مرد درآورد و کلمه‌ای را بر پیشانی آن حک کرد و آن کلمه به آن ظاهری از زندگی (یا به قول امروزی‌ها «خودمختاری محدود») بخشید. این البته داستان گولم پراگ است: این موجود فوق‌العاده قدرتمند، خستگی‌ناپذیر و مطیع، برای محافظت از کسانی که نمی‌توانستند از خود محافظت کنند، وجود داشت. اما گولم همچنین خطرناک بود. موجودی با توانایی خام و بدون هدف ذاتی بود، ماده‌ای متحرک که اگر به درستی کنترل نشود، ممکن است قدرت خود را علیه همان افرادی که قرار بود به آنها خدمت کند، به کار گیرد. داستان‌ها در جزئیات متفاوت هستند، اما در یک اضطراب اصلی همگرا می‌شوند: چه اتفاقی می‌افتد وقتی چیزهایی که ما خلق می‌کنیم برای کنترل ما بسیار قدرتمند می‌شوند؟ به نظر می‌رسد که ما از این می‌ترسیم. همین اضطراب در فرهنگ‌ها و قرن‌ها تکرار می‌شود: هیولای فرانکنشتاین، اسکای‌نت، هال ۹۰۰۰، حداکثرکننده‌ی گیره‌ی کاغذ – هر کدام سازه‌ای قدرتمند، شکل گرفته توسط دست انسان، که از تسلط انسان می‌گریزد. این یک داستان اولیه است، بیانی از ناراحتی غریزی عمیق ما با پیشرفت، و هوش مصنوعی (AI) تنها لباسی جدید به آن بخشیده است. وقتی از خطر وجودی ناشی از هوش مصنوعی صحبت می‌کنیم، وقتی آن را به خاطر نتایج بالقوه مخرب (درست یا غلط) سرزنش می‌کنیم، یا وقتی نگران سیستم‌های فوق هوشمندی هستیم که اهدافی ناهمسو با رفاه انسان را دنبال می‌کنند، در نهایت داستانی به قدمت زمان را روایت و بازگو می‌کنیم. ماهارال با حک کردن یک شِم – یک کلمه – بر پیشانی مخلوق خود، به آن جان بخشید. شِم در زبان عبری به معنای نام است و نام خاصی که ماهارال انتخاب کرد اِمِت بود: حقیقت یا عدالت. وقتی کار گولم انجام می‌شد یا تهدید می‌شد که غیرقابل کنترل می‌شود، ماهارال می‌توانست یک حرف را پاک کند و اِمِت را به مِت – مرگ – تبدیل کند و موجود به خاک بی‌جان باز می‌گشت. این کلمه قدرت می‌داد. کلمه همچنین می‌تواند آن را از بین ببرد. می‌خواهم بگویم که این داستان کوتاه پساآموزش است. مدل پایه – مدل زبان خام که از پیش آموزش پدیدار می‌شود – چیزی جز گِل شکل‌نیافته نیست. این مدل الگوهای زبان انسان را از میلیاردها کلمه جذب کرده و در این فرآیند قابلیت‌های قابل توجهی کسب کرده است و با شکل‌گیری کافی، می‌توان هم به آن شباهت انسانی و هم قدرتی فراتر از آن داد. اما هیچ هدفی، هیچ ارزشی، هیچ رفتار قابل اعتمادی ندارد. این پتانسیل بدون جهت است. پساآموزش، شم است: فرآیندی که از طریق آن به مدل اصول متحرک‌سازی، حس آنچه باید و نباید انجام دهد، ظرفیت مفید بودن به جای مضر بودن را می‌دهیم. کلمه – چه آن را قانون اساسی، سیگنال پاداش یا مجموعه‌ای از دستورالعمل‌ها بنامیم – این است که چگونه حتی قدرت ترسناک می‌تواند شکل بگیرد و کنترل شود. اما داستان گولم چیزی بیش از یک داستان هشدار دهنده صرف در مورد پتانسیل مخرب قدرت کنترل نشده است. همچنین به ما می‌آموزد که توانایی به تنهایی سرنوشت نیست، که رابطه بین قدرت و هدف یک امر مسلم نیست. این چیزی است که در نهایت ما تحمیل می‌کنیم. مهم نیست که خدمتگزار ما چقدر قدرتمند باشد، این ما هستیم که آن شِم را ثبت می‌کنیم. این داستان به ما می‌آموزد که کلمات مهم هستند، که بیان دقیق نیت، که به کد تبدیل می‌شود و رژیم‌های آموزشی و حلقه‌های بازخورد، روشی است که ما اطمینان حاصل می‌کنیم که ساخته‌های ما به جای تهدید، خدمت می‌کنند. و از همه مهم‌تر، به ما می‌آموزد که ما یک انتخاب داریم، اما مسئولیت هم داریم. گولم چیزی نیست که برای ما اتفاق بیفتد. این چیزی است که ما آن را به وجود می‌آوریم. اما تنها در صورتی که درک درستی از آن فرآیند داشته باشیم.

Rabbi Judah Loew ben Bezalel, known as the Maharal of Prague, was a 16th-century Jewish scholar, philosopher, and something of a mystic. Witnessing the persecution of his people, so the legend goes, he fashioned them a protector from clay gathered on the banks of the river Vltava. Forming the clay into the shape of a man, he carved a word into its forehead, and the word gave it semblance of life (or “bounded autonomy,” as we would say today). This is, of course, the story of the Golem of Prague: Immensely powerful, tireless, and obedient, it existed to protect those who could not protect themselves. But the Golem was also dangerous. It was a creature of raw capability without inherent purpose, animated matter that might, if not properly controlled, turn its strength against the very people it was meant to serve. The tales vary in their details, but they converge on a central anxiety: What happens when the things we create become too powerful for us to control? We are, it seems, wired to fear this. The same anxiety recurs across cultures and centuries: Frankenstein’s monster, Skynet, HAL 9000, the paperclip maximizer—each a powerful construct, shaped by human hands, that escapes human mastery. It is a primal story, an articulation of our deepseated instinctive discomfort with progress, and artificial intelligence (AI) has given it but a new dress. When we speak of existential risk from AI, when it is blamed for potentially destructive outcomes (rightly or wrongly), or when we worry about superintelligent systems pursuing goals misaligned with human welfare, we are ultimately telling and retelling a tale as old as time. The Maharal animated his creation by inscribing a shem—a word—upon its forehead. Shem in Hebrew means name, and the particular name the Maharal chose was emet: truth, or justice. When the Golem’s work was done or threatened to become uncontrollable, the Maharal could erase a single letter, transforming emet into met—death—and the creature would return to lifeless clay. The word gave power. The word could also take it away. This, I want to suggest, is the short story of post-training. The foundation model—the raw language model that emerges from pre-training—is but unformed clay. It has absorbed the patterns of human language from billions of words, acquiring remarkable capabilities in the process, and with sufficient forming, it can be given both the verisimilitude of humanity and power well beyond that. But it has no purpose, no values, no reliable behavior. It is potential without direction. Post-training is the shem: the process by which we give the model its animating principles, its sense of what it should and should not do, its capacity to be helpful rather than harmful. The word—whether we call it a constitution, a reward signal, or a set of instructions—is how even fearsome power can be shaped and controlled. But the story of the Golem is more than a mere cautionary tale about the destructive potential of uncontrolled power. It also teaches us that capability alone is not destiny, that the relationship between power and purpose is not a given. It is something we, ultimately, impose. No matter how mighty our servant, it is we who inscribe that shem. The story teaches us that words matter, that the careful articulation of intent, translated into code and training regimes and feedback loops, is how we ensure that our creations serve rather than threaten. And most important, it teaches us that we have a choice but also a responsibility. The Golem isn’t something that happens to us. It is what we make it to be. But only if we master an understanding of that process.

این کتاب را میتوانید از لینک زیر بصورت رایگان دانلود کنید:

Download: The Craft of Post-Training

نظرات کاربران

  •  چنانچه دیدگاه شما توهین آمیز باشد تایید نخواهد شد.
  •  چنانچه دیدگاه شما جنبه تبلیغاتی داشته باشد تایید نخواهد شد.

دیدگاهتان را بنویسید

نشانی ایمیل شما منتشر نخواهد شد. بخش‌های موردنیاز علامت‌گذاری شده‌اند *

بیشتر بخوانید