{"product_id":"ai-model-evaluation","title":"AI Model Evaluation","description":"\u003cp\u003eBefore you trust critical business systems to an AI model, you need to answer a few questions. Will it be fast enough? Will the system satisfy user expectations? Is it safe and can you trust the output? This book will help you answer these questions before you roll out an AI system, and make sure it runs smoothly after you deploy. \u003c\/p\u003e \u003cul\u003e  \u003cli data-leveltext=\"?\" data-font=\"Symbol\" data-listid=\"5\" data-list-defn-props='{\"335552541\":1,\"335559685\":720,\"335559991\":360,\"469769226\":\"Symbol\",\"469769242\":[8226],\"469777803\":\"left\",\"469777804\":\"?\",\"469777815\":\"multilevel\"}' data-aria-posinset=\"1\" data-aria-level=\"1\"\u003eLearn simple ways to test how your model behaves before it is used in real systems. \u003c\/li\u003e \u003c\/ul\u003e \u003cul\u003e  \u003cli data-leveltext=\"?\" data-font=\"Symbol\" data-listid=\"5\" data-list-defn-props='{\"335552541\":1,\"335559685\":720,\"335559991\":360,\"469769226\":\"Symbol\",\"469769242\":[8226],\"469777803\":\"left\",\"469777804\":\"?\",\"469777815\":\"multilevel\"}' data-aria-posinset=\"2\" data-aria-level=\"1\"\u003eTry your model with real data to see how it performs in real situations. \u003c\/li\u003e \u003c\/ul\u003e \u003cul\u003e  \u003cli data-leveltext=\"?\" data-font=\"Symbol\" data-listid=\"5\" data-list-defn-props='{\"335552541\":1,\"335559685\":720,\"335559991\":360,\"469769226\":\"Symbol\",\"469769242\":[8226],\"469777803\":\"left\",\"469777804\":\"?\",\"469777815\":\"multilevel\"}' data-aria-posinset=\"3\" data-aria-level=\"1\"\u003eDesign A\/B tests that validate model impact on key product metrics. \u003c\/li\u003e \u003c\/ul\u003e \u003cul\u003e  \u003cli data-leveltext=\"?\" data-font=\"Symbol\" data-listid=\"5\" data-list-defn-props='{\"335552541\":1,\"335559685\":720,\"335559991\":360,\"469769226\":\"Symbol\",\"469769242\":[8226],\"469777803\":\"left\",\"469777804\":\"?\",\"469777815\":\"multilevel\"}' data-aria-posinset=\"4\" data-aria-level=\"1\"\u003eSpot nuanced failures with human-in-the-loop feedback and qualitative evaluations. \u003c\/li\u003e \u003c\/ul\u003e \u003cul\u003e  \u003cli data-leveltext=\"?\" data-font=\"Symbol\" data-listid=\"5\" data-list-defn-props='{\"335552541\":1,\"335559685\":720,\"335559991\":360,\"469769226\":\"Symbol\",\"469769242\":[8226],\"469777803\":\"left\",\"469777804\":\"?\",\"469777815\":\"multilevel\"}' data-aria-posinset=\"5\" data-aria-level=\"1\"\u003eUse LLMs to help review and test models more quickly. \u003c\/li\u003e \u003c\/ul\u003e \u003cp\u003e\u003cstrong\u003eAI Model Evaluation?\u003c\/strong\u003eteaches you how to effectively evaluate and assess machine learning models for better scaling and integration. Each chapter looks at a different way to test a model, starting with offline evaluations and moving into live A\/B tests, shadow traffic deployments and LLM-based feedback loops. The book uses a hands-on example grounded in a movie recommendation engine. \u003c\/p\u003e \u003cp\u003eAfter reading this book, you will be able to evaluate both model behaviour and engineering system performance. You will have the tools to ensure your AI systems are effective and reliable in production. This book is for practitioners with experience in machine learning, data science, or software engineering. \u003c\/p\u003e","brand":"Gardners","offers":[{"title":"Default Title","offer_id":57504886194549,"sku":"9781633435674","price":45.99,"currency_code":"GBP","in_stock":true}],"thumbnail_url":"\/\/cdn.shopify.com\/s\/files\/1\/0612\/7193\/3106\/files\/9781633435674.jpg?v=1786953803","url":"https:\/\/backstory.london\/products\/ai-model-evaluation","provider":"Backstory","version":"1.0","type":"link"}