- عنوان کتاب: The Craft of Post-Training -A Practical Guide for AI Engineers and Developers
- نویسنده: Von Csefalvay, Chris
- حوزه: توسعه هوش مصنوعی
- سال انتشار: 2026
- تعداد صفحه: 434
- زبان اصلی: انگلیسی
- نوع فایل: pdf
- حجم فایل: 17.1 مگابایت
خاخام یهودا لو بن بزالل، معروف به مهارال پراگ، یک محقق، فیلسوف و تا حدودی عارف یهودی قرن شانزدهم بود. طبق افسانهها، او با مشاهده آزار و اذیت قوم خود، از گل رس جمعآوریشده در سواحل رودخانه ولتاوا، محافظی برای آنها ساخت. او گل را به شکل یک مرد درآورد و کلمهای را بر پیشانی آن حک کرد و آن کلمه به آن ظاهری از زندگی (یا به قول امروزیها «خودمختاری محدود») بخشید. این البته داستان گولم پراگ است: این موجود فوقالعاده قدرتمند، خستگیناپذیر و مطیع، برای محافظت از کسانی که نمیتوانستند از خود محافظت کنند، وجود داشت. اما گولم همچنین خطرناک بود. موجودی با توانایی خام و بدون هدف ذاتی بود، مادهای متحرک که اگر به درستی کنترل نشود، ممکن است قدرت خود را علیه همان افرادی که قرار بود به آنها خدمت کند، به کار گیرد. داستانها در جزئیات متفاوت هستند، اما در یک اضطراب اصلی همگرا میشوند: چه اتفاقی میافتد وقتی چیزهایی که ما خلق میکنیم برای کنترل ما بسیار قدرتمند میشوند؟ به نظر میرسد که ما از این میترسیم. همین اضطراب در فرهنگها و قرنها تکرار میشود: هیولای فرانکنشتاین، اسکاینت، هال ۹۰۰۰، حداکثرکنندهی گیرهی کاغذ – هر کدام سازهای قدرتمند، شکل گرفته توسط دست انسان، که از تسلط انسان میگریزد. این یک داستان اولیه است، بیانی از ناراحتی غریزی عمیق ما با پیشرفت، و هوش مصنوعی (AI) تنها لباسی جدید به آن بخشیده است. وقتی از خطر وجودی ناشی از هوش مصنوعی صحبت میکنیم، وقتی آن را به خاطر نتایج بالقوه مخرب (درست یا غلط) سرزنش میکنیم، یا وقتی نگران سیستمهای فوق هوشمندی هستیم که اهدافی ناهمسو با رفاه انسان را دنبال میکنند، در نهایت داستانی به قدمت زمان را روایت و بازگو میکنیم. ماهارال با حک کردن یک شِم – یک کلمه – بر پیشانی مخلوق خود، به آن جان بخشید. شِم در زبان عبری به معنای نام است و نام خاصی که ماهارال انتخاب کرد اِمِت بود: حقیقت یا عدالت. وقتی کار گولم انجام میشد یا تهدید میشد که غیرقابل کنترل میشود، ماهارال میتوانست یک حرف را پاک کند و اِمِت را به مِت – مرگ – تبدیل کند و موجود به خاک بیجان باز میگشت. این کلمه قدرت میداد. کلمه همچنین میتواند آن را از بین ببرد. میخواهم بگویم که این داستان کوتاه پساآموزش است. مدل پایه – مدل زبان خام که از پیش آموزش پدیدار میشود – چیزی جز گِل شکلنیافته نیست. این مدل الگوهای زبان انسان را از میلیاردها کلمه جذب کرده و در این فرآیند قابلیتهای قابل توجهی کسب کرده است و با شکلگیری کافی، میتوان هم به آن شباهت انسانی و هم قدرتی فراتر از آن داد. اما هیچ هدفی، هیچ ارزشی، هیچ رفتار قابل اعتمادی ندارد. این پتانسیل بدون جهت است. پساآموزش، شم است: فرآیندی که از طریق آن به مدل اصول متحرکسازی، حس آنچه باید و نباید انجام دهد، ظرفیت مفید بودن به جای مضر بودن را میدهیم. کلمه – چه آن را قانون اساسی، سیگنال پاداش یا مجموعهای از دستورالعملها بنامیم – این است که چگونه حتی قدرت ترسناک میتواند شکل بگیرد و کنترل شود. اما داستان گولم چیزی بیش از یک داستان هشدار دهنده صرف در مورد پتانسیل مخرب قدرت کنترل نشده است. همچنین به ما میآموزد که توانایی به تنهایی سرنوشت نیست، که رابطه بین قدرت و هدف یک امر مسلم نیست. این چیزی است که در نهایت ما تحمیل میکنیم. مهم نیست که خدمتگزار ما چقدر قدرتمند باشد، این ما هستیم که آن شِم را ثبت میکنیم. این داستان به ما میآموزد که کلمات مهم هستند، که بیان دقیق نیت، که به کد تبدیل میشود و رژیمهای آموزشی و حلقههای بازخورد، روشی است که ما اطمینان حاصل میکنیم که ساختههای ما به جای تهدید، خدمت میکنند. و از همه مهمتر، به ما میآموزد که ما یک انتخاب داریم، اما مسئولیت هم داریم. گولم چیزی نیست که برای ما اتفاق بیفتد. این چیزی است که ما آن را به وجود میآوریم. اما تنها در صورتی که درک درستی از آن فرآیند داشته باشیم.
Rabbi Judah Loew ben Bezalel, known as the Maharal of Prague, was a 16th-century Jewish scholar, philosopher, and something of a mystic. Witnessing the persecution of his people, so the legend goes, he fashioned them a protector from clay gathered on the banks of the river Vltava. Forming the clay into the shape of a man, he carved a word into its forehead, and the word gave it semblance of life (or “bounded autonomy,” as we would say today). This is, of course, the story of the Golem of Prague: Immensely powerful, tireless, and obedient, it existed to protect those who could not protect themselves. But the Golem was also dangerous. It was a creature of raw capability without inherent purpose, animated matter that might, if not properly controlled, turn its strength against the very people it was meant to serve. The tales vary in their details, but they converge on a central anxiety: What happens when the things we create become too powerful for us to control? We are, it seems, wired to fear this. The same anxiety recurs across cultures and centuries: Frankenstein’s monster, Skynet, HAL 9000, the paperclip maximizer—each a powerful construct, shaped by human hands, that escapes human mastery. It is a primal story, an articulation of our deepseated instinctive discomfort with progress, and artificial intelligence (AI) has given it but a new dress. When we speak of existential risk from AI, when it is blamed for potentially destructive outcomes (rightly or wrongly), or when we worry about superintelligent systems pursuing goals misaligned with human welfare, we are ultimately telling and retelling a tale as old as time. The Maharal animated his creation by inscribing a shem—a word—upon its forehead. Shem in Hebrew means name, and the particular name the Maharal chose was emet: truth, or justice. When the Golem’s work was done or threatened to become uncontrollable, the Maharal could erase a single letter, transforming emet into met—death—and the creature would return to lifeless clay. The word gave power. The word could also take it away. This, I want to suggest, is the short story of post-training. The foundation model—the raw language model that emerges from pre-training—is but unformed clay. It has absorbed the patterns of human language from billions of words, acquiring remarkable capabilities in the process, and with sufficient forming, it can be given both the verisimilitude of humanity and power well beyond that. But it has no purpose, no values, no reliable behavior. It is potential without direction. Post-training is the shem: the process by which we give the model its animating principles, its sense of what it should and should not do, its capacity to be helpful rather than harmful. The word—whether we call it a constitution, a reward signal, or a set of instructions—is how even fearsome power can be shaped and controlled. But the story of the Golem is more than a mere cautionary tale about the destructive potential of uncontrolled power. It also teaches us that capability alone is not destiny, that the relationship between power and purpose is not a given. It is something we, ultimately, impose. No matter how mighty our servant, it is we who inscribe that shem. The story teaches us that words matter, that the careful articulation of intent, translated into code and training regimes and feedback loops, is how we ensure that our creations serve rather than threaten. And most important, it teaches us that we have a choice but also a responsibility. The Golem isn’t something that happens to us. It is what we make it to be. But only if we master an understanding of that process.
این کتاب را میتوانید از لینک زیر بصورت رایگان دانلود کنید:
Download: The Craft of Post-Training





نظرات کاربران