194 Best Computer Vision Books of All Time
We've ranked the best computer vision books using expert recommendations, sales data, and millions of reader ratings. At Shortform, we know books. Our book guides are the best in the world. Learn why.
1
2Code: The Hidden Language of Computer Hardware and Software
Using everyday objects and familiar language systems such as Braille and Morse code, author Charles Petzold weaves an illuminating narrative for anyone who’s ever wondered about the secret inner life of computers and other smart machines.
It’s a cleverly illustrated and eminently comprehensible story—and along the way, you’ll discover you’ve gained a real context for understanding today’s world of PCs, digital media, and the Internet. No matter what your level of technical savvy, CODE will charm you—and perhaps even awaken the technophile within.
It gets you to use your imagination to virtually build a computer. It’s easy to read, you can lie down on the couch and enjoy it—it’s not so much of a textbook. It demystifies the magic of a computer and what it is. [source]
Hands-On Machine Learning with Scikit-Learn and TensorFlow: Concepts, Tools, and Techniques to Build Intelligent Systems
By using concrete examples, minimal theory, and two production-ready Python frameworks-scikit-learn and TensorFlow-author Aurélien Géron helps you gain an intuitive understanding of the concepts and tools for building intelligent systems. You'll learn a range of techniques, starting with simple linear regression and progressing to deep neural networks. With exercises in each chapter to help you apply what you've learned, all you need is programming experience to get started.
Explore the machine learning landscape, particularly neural nets Use scikit-learn to track an example machine-learning project end-to-end Explore several training models, including support vector machines, decision trees, random forests, and ensemble methods Use the TensorFlow library to build and train neural nets Dive into neural net architectures, including convolutional nets, recurrent nets, and deep reinforcement learning Learn techniques for training and scaling deep neural nets Apply practical code examples without acquiring excessive machine learning theory or algorithm detailsBook to Start You on Machine Learning - KDnuggets https://t.co/19fdX59b0d This book is “Hands-On Machine Learning with Scikit-Learn & TensorFlow”. each new revision has become an even better version of one of the best in-depth resources to learn Machine Learning by doing. https://t.co/ujyUH3xU3e [source]
4Hands-On Machine Learning with Scikit-Learn, Keras, and Tensorflow: Concepts, Tools, and Techniques to Build Intelligent Systems
Calvin and Hobbes (Calvin and Hobbes #1)
6Design Patterns: Elements of Reusable Object-Oriented Software
The authors begin by describing what patterns are and how they can help you design object-oriented software. They then go on to systematically name, explain, evaluate, and catalog recurring designs in object-oriented systems. With Design Patterns as your guide, you will learn how these important patterns fit into the software development process, and how you can leverage them to solve your own design problems most efficiently.
Each pattern describes the circumstances in which it is applicable, when it can be applied in view of other design constraints, and the consequences and trade-offs of using the pattern within a larger design. All patterns are compiled from real systems and are based on real-world examples. Each pattern also includes code that demonstrates how it may be implemented in object-oriented programming languages like C++ or Smalltalk.
7Nexus: A Brief History of Information Networks from the Stone Age to AI
8Programming Computer Vision with Python: Tools and algorithms for analyzing images
Programming Computer Vision with Python explains computer vision in broad terms that won’t bog you down in theory. You get complete code samples with explanations on how to reproduce and build upon each example, along with exercises to help you apply what you’ve learned. This book is ideal for students, researchers, and enthusiasts with basic programming and standard mathematical skills.
Learn techniques used in robot navigation, medical image analysis, and other computer vision applications
Work with image mappings and transforms, such as texture warping and panorama creation
Compute 3D reconstructions from several images of the same scene
Organize images based on similarity or content, using clustering methods
Build efficient image retrieval techniques to search for images based on visual content
Use algorithms to classify image content and recognize objects
Access the popular OpenCV library through a Python interface
Computer Vision: Models, Learning, and Inference
10Tinyml: Machine Learning with Tensorflow Lite on Arduino and Ultra-Low-Power Microcontrollers
Authors Pete Warden and Daniel Situnayake explain how you can train models that are small enough to fit into any environment, including small embedded devices that can run for a year or more on a single coin cell battery. Ideal for software and hardware developers who want to build embedded devices using machine learning, this guide shows you how to create a TinyML project step-by-step. No machine learning or microcontroller experience is necessary.
Learn practical machine learning applications on embedded devices, including simple uses such as speech recognition and gesture detection
Train models such as speech, accelerometer, and image recognition, you can deploy on Arduino and other embedded platforms
Understand how to work with Arduino and ultralow-power microcontrollers
Use techniques for optimizing latency, energy usage, and model and binary size
11Why Machines Learn: The Elegant Math Behind Modern AI
Building Machine Learning Powered Applications: Going from Idea to Product
Author Emmanuel Ameisen, who worked as a data scientist at Zipcar and led Insight Data Science's AI program, demonstrates key ML concepts with code snippets, illustrations, and screenshots from the book's example application.
The first part of this guide shows you how to plan and measure success for an ML application. Part II shows you how to build a working ML model, and Part III explains how to improve the model until it fulfills your original vision. Part IV covers deployment and monitoring strategies.
This book will help you:
Determine your product goal and set up a machine learning problem
Build your first end-to-end pipeline quickly and acquire an initial dataset
Train and evaluate your ML model and address performance bottlenecks
Deploy and monitor models in a production environment
13Machine Learning: A Probabilistic Perspective
Today's Web-enabled deluge of electronic data calls for automated methods of data analysis. Machine learning provides these, developing methods that can automatically detect patterns in data and then use the uncovered patterns to predict future data. This textbook offers a comprehensive and self-contained introduction to the field of machine learning, based on a unified, probabilistic approach.
The coverage combines breadth and depth, offering necessary background material on such topics as probability, optimization, and linear algebra as well as discussion of recent developments in the field, including conditional random fields, L1 regularization, and deep learning. The book is written in an informal, accessible style, complete with pseudo-code for the most important algorithms. All topics are copiously illustrated with color images and worked examples drawn from such application domains as biology, text processing, computer vision, and robotics. Rather than providing a cookbook of different heuristic methods, the book stresses a principled model-based approach, often using the language of graphical models to specify models in a concise and intuitive way. Almost all the models described have been implemented in a MATLAB software package—PMTK (probabilistic modeling toolkit)—that is freely available online. The book is suitable for upper-level undergraduates with an introductory-level college math background and beginning graduate students.
#MachineLearning — a Probabilistic Perspective: https://t.co/wAZwLoUFGF ———— #BigData #Statistics #DataScience #DeepLearning #AI #Algorithms #StatisticalLiteracy #Mathematics #abdsc ——— ⬇Get this brilliant 1100-page 28-chapter highly-rated book: https://t.co/Tm2zchpHSu https://t.co/jprUDdzkj8 [source]
14Practical Deep Learning for Cloud, Mobile, and Edge: Real-World AI & Computer-Vision Projects Using Python, Keras & Tensorflow
Relying on years of industry experience transforming deep learning research into award-winning applications, Anirudh Koul, Siddha Ganju, and Meher Kasam guide you through the process of converting an idea into something that people in the real world can use.
Train, tune, and deploy computer vision models with Keras, TensorFlow, Core ML, and TensorFlow Lite
Develop AI for a range of devices including Raspberry Pi, Jetson Nano, and Google Coral
Explore fun projects, from Silicon Valley's Not Hotdog app to 40+ industry case studies
Simulate an autonomous car in a video game environment and build a miniature version with reinforcement learning
Use transfer learning to train models in minutes
Discover 50+ practical tips for maximizing model accuracy and speed, debugging, and scaling to millions of users
15Pattern Classification
An Instructor's Manual presenting detailed solutions to all the problems in the book is available from the Wiley editorial department.
Managing Director/Thiel Capital
Eric Weinstein recommended this book on Twitter. [source]
PyTorch Computer Vision Cookbook: Over 70 recipes to master the art of computer vision with deep learning and PyTorch 1.x
17The Data Science Design Manual
In particular, the book stresses the following basic principles as fundamental to becoming a good data scientist: "Valuing Doing the Simple Things Right," laying the groundwork of what really matters in analyzing data; "Developing Mathematical Intuition," so that readers can understand on an intuitive level why these concepts were developed, how they are useful and when they work best, and; "Thinking Like a Computer Scientist, but Acting Like a Statistician," following approaches which come most naturally to computer scientists while maintaining the core values of statistical reasoning. The book does not emphasize any particular language or suite of data analysis tools, but instead provides a high-level discussion of important design principles.
This book covers enough material for an "Introduction to Data Science" course at the undergraduate or early graduate student levels. A full set of lecture slides for teaching this course are available at an associated website, along with data resources for projects and assignments, and online video lectures.
Other Pedagogical features of this book include: "War Stories" offering perspectives on how data science techniques apply in the real world; "False Starts" revealing the subtle reasons why certain approaches fail; "Take-Home Lessons" emphasizing the big-picture concepts to learn from each chapter; "Homework Problems" providing a wide range of exercises for self-study; "Kaggle Challenges" from the online platform Kaggle; examples taken from the data science television show "The Quant Shop," and; concluding notes in each tutorial chapter pointing readers to primary sources and additional references.
18Learning OpenCV: Computer Vision with the OpenCV Library
-William T. Freeman, Computer Science and Artificial Intelligence Laboratory, Massachusetts Institute of Technology
Learning OpenCV puts you in the middle of the rapidly expanding field of computer vision. Written by the creators of the free open source OpenCV library, this book introduces you to computer vision and demonstrates how you can quickly build applications that enable computers to "see" and make decisions based on that data.
Computer vision is everywhere-in security systems, manufacturing inspection systems, medical image analysis, Unmanned Aerial Vehicles, and more. It stitches Google maps and Google Earth together, checks the pixels on LCD screens, and makes sure the stitches in your shirt are sewn properly. OpenCV provides an easy-to-use computer vision framework and a comprehensive library with more than 500 functions that can run vision code in real time.
Learning OpenCV will teach any developer or hobbyist to use the framework quickly with the help of hands-on exercises in each chapter. This book includes:
A thorough introduction to OpenCV Getting input from cameras Transforming images Segmenting images and shape matching Pattern recognition, including face detection Tracking and motion in 2 and 3 dimensions 3D reconstruction from stereo vision Machine learning algorithms Getting machines to see is a challenging but entertaining goal. Whether you want to build simple or sophisticated vision applications, Learning OpenCV is the book you need to get started.
Computer Vision: Algorithms and Applications
Computer Vision: Algorithms and Applications explores the variety of techniques commonly used to analyze and interpret images. It also describes challenging real-world applications where vision is being successfully used, both for specialized applications such as medical imaging, and for fun, consumer-level tasks such as image editing and stitching, which students can apply to their own personal photos and videos.
More than just a source of "recipes," this exceptionally authoritative and comprehensive textbook/reference also takes a scientific approach to basic vision problems, formulating physical models of the imaging process before inverting them to produce descriptions of a scene. These problems are also analyzed using statistical models and solved using rigorous engineering techniques
Topics and features:
Structured to support active curricula and project-oriented courses, with tips in the Introduction for using the book in a variety of customized courses Presents exercises at the end of each chapter with a heavy emphasis on testing algorithms and containing numerous suggestions for small mid-term projects Provides additional material and more detailed mathematical topics in the Appendices, which cover linear algebra, numerical techniques, and Bayesian estimation theory Suggests additional reading at the end of each chapter, including the latest research in each sub-field, in addition to a full Bibliography at the end of the book Supplies supplementary course material for students at the associated website, http: //szeliski.org/Book/ Suitable for an upper-level undergraduate or graduate-level course in computer science or engineering, this textbook focuses on basic techniques that work under real-world conditions and encourages students to push their creative boundaries. Its design and exposition also make it eminently suitable as a unique reference to the fundamental techniques and current research literature in computer vision.
20Learning From Data: A Short Course
21Hands-On Machine Learning with Scikit-Learn and PyTorch: Concepts, Tools, and Techniques to Build Intelligent Systems
The potential of machine learning today is extraordinary, yet many aspiring developers and tech professionals find themselves daunted by its complexity. Whether you're looking to enhance your skill set and apply machine learning to real-world projects or are simply curious about how AI systems function, this book is your jumping-off place.With an approachable yet deeply informative style, author Aurélien Géron delivers the ultimate introductory guide to machine learning and deep learning. Drawing on the Hugging Face ecosystem, with a focus on clear explanations and real-world examples, the book takes you through cutting-edge tools like Scikit-Learn and PyTorch—from basic regression techniques to advanced neural networks. Whether you're a student, professional, or hobbyist, you'll gain the skills to build intelligent systems.Understand ML basics, including concepts like overfitting and hyperparameter tuningComplete an end-to-end ML project using scikit-Learn, covering everything from data exploration to model evaluationLearn techniques for unsupervised learning, such as clustering and anomaly detectionBuild advanced architectures like transformers and diffusion models with PyTorchHarness the power of pretrained models—including LLMs—and learn to fine-tune themTrain autonomous agents using reinforcement learning
22Vision Language Models: Building VLMs with Hugging Face
Vision language models (VLMs) combine computer vision and natural language processing to create powerful systems that can interpret, generate, and respond in multimodal contexts. Vision Language Models is a hands-on guide to building real-world VLMs using the most up-to-date stack of machine learning tools from Hugging Face, Meta (PyTorch), NVIDIA (Cuda), and others, written by leading researchers and practitioners Merve Noyan, Miquel Farré, Andrés Marafioti, and Orr Zohar. From image captioning and document understanding to advanced zero-shot inference and retrieval-augmented generation (RAG), this book covers the full VLM application and development lifecycle.Designed for ML engineers, data scientists, and developers, this guide distills cutting-edge VLM research into practical techniques. Readers will learn how to prepare datasets, select the right architectures, fine-tune and deploy models, and apply them to real-world tasks across a range of industries.Explore core model architectures and alignment techniquesTrain and fine-tune VLMs with Hugging Face, PyTorch, and othersDeploy models for applications like image search and captioningImplement advanced inference strategies, from zero-shot to agentic systemsBuild scalable VLM systems ready for production use
23Sutskever's List: Foundational ideas of modern AI
Get the eBook free when you register your print book at Manning."A perspective the field has needed. Sutskever’s List delivers it with care and historical accuracy.”—Yanping Huang, GoogleSutskever’s List is a guided intellectual journey through the ideas that made modern AI suddenly possible. Each chapter is anchored in specific papers, books, or other sources from Sutskever’s list. The papers themselves are not the focus. Instead, the author uses them as entry points into the larger breakthroughs, arguments, interconnections, and shifts in thinking that transformed the field.It begins with AlexNet, where data, GPUs, and training craft made neural networks impossible to dismiss, then moves to ResNet, where depth becomes a superpower rather than a liability. From there, the story accelerates through sequence models, speech systems, attention, Transformers, and hyperscale, showing how AI escaped older bottlenecks and became built to grow.Later chapters ask whether these systems can reason, why simplicity can emerge from complexity, and what intelligence and safety mean once AI capabilities begin to feel uncanny. Reviewers praise Heimann’s “exquisitely deep, detailed, and nuanced knowledge” and the “massive amount of gold material” gathered here. Yet the book remains remarkably easy to read, turning difficult papers into a “guided initiation those papers were never designed to provide on their own.”As you go, you’ll understand how abstract lab results have translated into real-world consequences, including shifting architectures and internal organizational politics. With lucid explanations of the core technologies of AI as defined in Sutskever’s collection of seminal papers, Heimann explores common engineering choices, evaluating the strengths and limits of deep learning without falling for hype or cynicism. Complex concepts are clarified through relevant examples, vivid anecdotes, and practical engineering insights.Each of the core papers examined in Sutskever’s List represents a crucial steppingstone in the evolution of the AI. You’ll love how Richard Heimann combines a deep technical background with a journalistic eye, never losing sight of practical considerations and providing a stepping off point to understand where the technology goes next.Sutskever’s List features nine chapters, an epilogue, and a practical appendix, smoothly blending technical instruction with cultural and historical context. The result is a logically flowing book that remains highly accessible, navigable, and technically deep without requiring the reader to have a specialist’s background.What's inside• Decoding landmark AI papers from AlexNet to transformers• Understanding scaling laws, reasoning models, and AI safety• Engineering patterns that scale from research to real-world systemsAbout the readerFor anyone interested in modern AI and deep learning. No specialist knowledge required.About the authorRichard Heimann has honed his deep AI and machine learning expertise across technical and strategic roles in industry, academia, and government. He excels at translating complex ideas into clear, engaging insights for audiences from practitioners to policymakers.Table of Contents1 What did Ilya see?2 The AlexNet moment3 ResNet revolution4 Deep learning accelerates5 Attention is all you need6 The birth of hyperscale7 The pivot to reasoning8 Simplicity, hidden in complexity9 Safe superintelligenceEpilogue: The missing piecesAppendix: Design patterns for engineers
24Nexus: Una breve historia de las redes de información desde la edad de piedra ha sta la IA / Nexus: A Brief History of Inform
Mapping and Visualization with Supercollider
26Information Theory, Inference and Learning Algorithms
Natural Image Statistics: A Probablistic Approach To Early Computational Vision. (Computational Imaging And Vision)
Autonomous Intelligent Vehicles: Theory, Algorithms, and Implementation
This important text/reference presents state-of-the-art research on intelligent vehicles, covering not only topics of object/obstacle detection and recognition, but also aspects of vehicle motion control. With an emphasis on both high-level concepts, and practical detail, the text links theory, algorithms, and issues of hardware and software implementation in intelligent vehicle research.
Topics and features: presents a thorough introduction to the development and latest progress in intelligent vehicle research, and proposes a basic framework; provides detection and tracking algorithms for structured and unstructured roads, as well as on-road vehicle detection and tracking algorithms using boosted Gabor features; discusses an approach for multiple sensor-based multiple-object tracking, in addition to an integrated DGPS/IMU positioning approach; examines a vehicle navigation approach using global views; introduces algorithms for lateral and longitudinal vehicle motion control.
An essential reference for researchers in the field, the broad coverage of all aspects of this research will also appeal to graduate students of computer science and robotics who are interested in intelligent vehicles.
29Generative Deep Learning: Teaching Machines to Paint, Write, Compose, and Play
With this practical book, machine learning engineers and data scientists will learn how to recreate some of the most famous examples of generative deep learning models, such as variational autoencoders and generative adversarial networks (GANs). You'll also learn how to apply the techniques to your own datasets.
David Foster, cofounder of Applied Data Science, demonstrates the inner workings of each technique, starting with the basics of deep learning before advancing to the most cutting-edge algorithms in the field. Through tips and tricks, you'll learn how to make your models learn more efficiently and become more creative.
Get a fundamental overview of deep learning
Learn about libraries such as Keras and TensorFlow
Discover how variational autoencoders work
Get practical examples of generative adversarial networks (GANs)
Understand how autoregressive generative models function
Apply generative models within a reinforcement learning setting to accomplish tasks
Applied Artificial Intelligence: A Handbook for Business Leaders
We teach you how to lead successful AI initiatives by prioritizing the right opportunities, building a diverse team of experts, conducting strategic experiments, and consciously designing your solutions to benefit both your organization and society as a whole.
Fantastic presentation on #ConversationalAI and #chatbots at #PegaWorld by @thinkmariya — check out: https://t.co/lOT50RlkWD @topbots #AI #MachineLearning #NLProc #NLU #NLG See also her book: https://t.co/pyWfEBnANU https://t.co/8q3uyDJXwU [source]
31AI: Unexplainable, Unpredictable, Uncontrollable
Delving into the deeply enigmatic nature of Artificial Intelligence (AI), AI: Unexplainable, Unpredictable, Uncontrollable explores the various reasons why the field is so challenging. Written by one of the founders of the field of AI safety, this book addresses some of the most fascinating questions facing humanity, including the nature of intelligence, consciousness, values, and knowledge.Moving from a broad introduction to the core problems, such as the unpredictability of AI outcomes or the difficulty in explaining AI decisions, this book arrives at more complex questions of ownership and control, conducting an in-depth analysis of potential hazards and unintentional consequences. The book then concludes with philosophical and existential considerations, probing into questions of AI personhood, consciousness, and the distinction between human intelligence and artificial general intelligence (AGI).Bridging the gap between technical intricacies and philosophical musings, AI: Unexplainable, Unpredictable, Uncontrollable appeals to both AI experts and enthusiasts looking for a comprehensive understanding of the field, while also being written for a general audience with minimal technical jargon.
32The AI Workshop: The Complete Beginner's Guide to AI: Your A-Z Guide to Mastering Artificial Intelligence for Life, Work, and
Computer Vision:: A Modern Approach
34Bandit Algorithms
35Foundations of Computer Vision (Adaptive Computation and Machine Learning series)
Computational Line Geometry
Optics: Learning by Computing, with Examples Using MathCAD
Couscous and Other Good Food from Morocco
OpenCV for Secret Agents
40Superagency: What Could Possibly Go Right with Our AI Future
41Speech and Language Processing: An Introduction to Natural Language Processing, Computational Linguistics and Speech Recognition
- UNIFIED AND COMPREHENSIVE COVERAGE OF THE FIELD
Covers the fundamental algorithms of each field, whether proposed for spoken or written language, whether logical or statistical in origin.
- EMPHASIS ON WEB AND OTHER PRACTICAL APPLICATIONS
Gives readers an understanding of how language-related algorithms can be applied to important real-world problems.
- EMPHASIS ON SCIENTIFIC EVALUATION
Offers a description of how systems are evaluated with each problem domain.
- EMPERICIST/STATISTICAL/MACHINE LEARNING APPROACHES TO LANGUAGE PROCESSING
Covers all the new statistical approaches, while still completely covering the earlier more structured and rule-based methods.
42Probabilistic Graphical Models: Principles and Techniques
Most tasks require a person or an automated system to reason—to reach conclusions based on available information. The framework of probabilistic graphical models, presented in this book, provides a general approach for this task. The approach is model-based, allowing interpretable models to be constructed and then manipulated by reasoning algorithms. These models can also be learned automatically from data, allowing the approach to be used in cases where manually constructing a model is difficult or even impossible. Because uncertainty is an inescapable aspect of most real-world applications, the book focuses on probabilistic models, which make the uncertainty explicit and provide models that are more faithful to reality.
Probabilistic Graphical Models discusses a variety of models, spanning Bayesian networks, undirected Markov networks, discrete and continuous models, and extensions to deal with dynamical systems and relational data. For each class of models, the text describes the three fundamental cornerstones: representation, inference, and learning, presenting both basic concepts and advanced techniques. Finally, the book considers the use of the proposed framework for causal reasoning and decision making under uncertainty. The main text in each chapter provides the detailed technical development of the key ideas. Most chapters also include boxes with additional material: skill boxes, which describe techniques; case study boxes, which discuss empirical cases related to the approach described in the text, including applications in computer vision, robotics, natural language understanding, and computational biology; and concept boxes, which present significant concepts drawn from the material in the chapter. Instructors (and readers) can group chapters in various combinations, from core topics to more technically advanced material, to suit their particular needs.
43Deep Thinking: Where Machine Intelligence Ends and Human Creativity Begins
That moment was more than a century in the making, and in this breakthrough book, Kasparov reveals his astonishing side of the story for the first time. He describes how it felt to strategize against an implacable, untiring opponent with the whole world watching, and recounts the history of machine intelligence through the microcosm of chess, considered by generations of scientific pioneers to be a key to unlocking the secrets of human and machine cognition. Kasparov uses his unrivaled experience to look into the future of intelligent machines and sees it bright with possibility. As many critics decry artificial intelligence as a menace, particularly to human jobs, Kasparov shows how humanity can rise to new heights with the help of our most extraordinary creations, rather than fear them. Deep Thinking is a tightly argued case for technological progress, from the man who stood at its precipice with his own career at stake.
Author
The great Garry Kasparov takes on the key economic issue of our time: how we can thrive as humans in a world of thinking machines. This important and optimistic book explains what we as humans are uniquely qualified to do. Instead or wringing our hands about robots, we should all read this book and embrace the future. [source]
Author
Garry Kasparov's perspectives on artificial intelligence are borne of personal experience - and despite that, are optimistic, wise and compelling. It's one thing for the giants of Silicon Valley to tell us our future is bright; it is another thing to hear it from the man who squared off with the world's most powerful computer, with the whole world watching, and his very identity at stake. [source]
Co-founder/PayPal, CEO/Affirm, Investor
A highly human exploration of artificial intelligence, its exciting possibilities and inherent limits. [source]
44Understanding Machine Learning: From Theory to Algorithms
45Linear Algebra for Data Science, Machine Learning, and Signal Processing
Maximise student engagement and understanding of matrix methods in data-driven applications with this modern teaching package. Students are introduced to matrices in two preliminary chapters, before progressing to advanced topics such as the nuclear norm, proximal operators and convex optimization. Highlighted applications include low-rank approximation, matrix completion, subspace learning, logistic regression for binary classification, robust PCA, dimensionality reduction and Procrustes problems. Extensively classroom-tested, the book includes over 200 multiple-choice questions suitable for in-class interactive learning or quizzes, as well as homework exercises (with solutions available for instructors). It encourages active learning with engaging 'explore' questions, with answers at the back of each chapter, and Julia code examples to demonstrate how the mathematics is actually used in practice. A suite of computational notebooks offers a hands-on learning experience for students. This is a perfect textbook for upper-level undergraduates and first-year graduate students who have taken a prior course in linear algebra basics.
Computer Vision with OpenCV 3 and Qt5: Build visually appealing, multithreaded, cross-platform computer vision applications
Inside PixInsight (The Patrick Moore Practical Astronomy Series)
48The Grammar of Graphics
GANs in Action: Deep learning with Generative Adversarial Networks
GANs in Action: Deep learning with Generative Adversarial Networks teaches you how to build and train your own generative adversarial networks. First, you'll get an introduction to generative modelling and how GANs work, along with an overview of their potential uses. Then, you'll start building your own simple adversarial system, as you explore the foundation of GAN architecture: the generator and discriminator networks.
Purchase of the print book includes a free eBook in PDF, Kindle, and ePub formats from Manning Publications.
Amazon Echo Show 8 User Guide: The Complete User Manual for Beginners and Pro to Master the New Amazon Echo Show 8 with Tips & Tricks for Alexa Skills
Overview of Echo Show 8
Setting up your Amazon Echo Show 8
Setup Alexa Voice Profiles
Setup Amazon Household & FreeTime
Customize the Home Screen on Your Echo Show
Add Amazon & Facebook Photos to Echo Show Home Screen
Set up Routines
Alexa Blueprint
Listen to Radio & Podcasts on Amazon Echo Show
Listen to Music on Amazon Echo Show
Listen to Audiobooks on Amazon Echo Show
Using Skype on Echo Show
Watch YouTube, Netflix & Amazon Prime Videos
Setup Smart Home Devices & Control your Appliances
Alexa Intercom, Drop-In, and Privacy
Phone Calls and Messaging
Setting up IFTTT
Get Weather & Traffic Updates
Flash Briefings
Reminder, Timers & Alarms
Alexa Skills, Questions & Eastern Eggs
Troubleshooting
And other Amazon Echo Show Settings
Don't wait, get this guide now by clicking the BUY NOW button and learn everything about your Echo Show 8!
51Linear Algebra and Learning from Data
52What Is ChatGPT Doing ... and Why Does It Work?
Mining of Massive Datasets
Augmented Mind: AI, Humans and the Superhuman Revolution
Hardware and Software Support for Virtualization
Despite the focus on architectural support in current architectures, some historical perspective is necessary to appropriately frame the problem. The first half of the book provides the historical perspective of the theoretical framework developed four decades ago by Popek and Goldberg. It also describes earlier systems that enabled virtualization despite the lack of architectural support in hardware.
As is often the case, theory defines a necessary-but not sufficient-set of features, and modern architectures are the result of the combination of the theoretical framework with insights derived from practical systems. The second half of the book describes state-of-the-art support for virtualization in both x86-64 and ARM processors. This book includes an in-depth description of the CPU, memory, and I/O virtualization of these two processor architectures, as well as case studies on the Linux/KVM, VMware, and Xen hypervisors. It concludes with a performance comparison of virtualization on current-generation x86- and ARM-based systems across multiple hypervisors.
56Dive into Deep Learning
Deep learning has revolutionized pattern recognition, introducing tools that power a wide range of technologies in such diverse fields as computer vision, natural language processing, and automatic speech recognition. Applying deep learning requires you to simultaneously understand how to cast a problem, the basic mathematics of modeling, the algorithms for fitting your models to data, and the engineering techniques to implement it all. This book is a comprehensive resource that makes deep learning approachable, while still providing sufficient technical depth to enable engineers, scientists, and students to use deep learning in their own work. No previous background in machine learning or deep learning is required―every concept is explained from scratch and the appendix provides a refresher on the mathematics needed. Runnable code is featured throughout, allowing you to develop your own intuition by putting key ideas into practice.
57Mastering OpenCV with Practical Computer Vision Projects
Deep Learning for Computer Vision with Python — Starter Bundle
Amazon Echo Dot - The Complete User Guide: Learn to Use Your Echo Dot Like A Pro
60The Perfect Bet: How Science and Math Are Taking the Luck Out of Gambling
61Learning OpenCV 3: Computer Vision in C++ with the OpenCV Library
With over 500 functions that span many areas in vision, OpenCV is used for commercial applications such as security, medical imaging, pattern and face recognition, robotics, and factory product inspection. This book gives you a firm grounding in computer vision and OpenCV for building simple or sophisticated vision applications. Hands-on exercises in each chapter help you apply what you've learned.
This volume covers the entire library, in its modern C++ implementation, including machine learning tools for computer vision.
Learn OpenCV data types, array types, and array operations
Capture and store still and video images with HighGUI
Transform images to stretch, shrink, warp, remap, and repair
Explore pattern recognition, including face detection
Track objects and motion through the visual field
Reconstruct 3D images from stereo vision
Discover basic and advanced machine learning techniques in OpenCV
62Foundations of Data Science
63Deep Learning for Finance
64AI Systems Performance Engineering: Optimizing Model Training and Inference Workloads with GPUs, CUDA, and PyTorch
Elevate your AI system performance capabilities with this definitive guide to maximizing efficiency across every layer of your AI infrastructure. In today's era of ever-growing generative models, AI Systems Performance Engineering provides engineers, researchers, and developers with a hands-on set of actionable optimization strategies. Learn to co-optimize hardware, software, and algorithms to build resilient, scalable, and cost-effective AI systems that excel in both training and inference. Authored by Chris Fregly, a performance-focused engineering and product leader, this resource transforms complex AI systems into streamlined, high-impact AI solutions.Inside, you'll discover step-by-step methodologies for fine-tuning GPU CUDA kernels, PyTorch-based algorithms, and multinode training and inference systems. You'll also master the art of scaling GPU clusters for high performance, distributed model training jobs, and inference servers. The book ends with a 175+-item checklist of proven, ready-to-use optimizations.Codesign and optimize hardware, software, and algorithms to achieve maximum throughput and cost savingsImplement cutting-edge inference strategies that reduce latency and boost throughput in real-world settingsUtilize industry-leading scalability tools and frameworksProfile, diagnose, and eliminate performance bottlenecks across complex AI pipelinesIntegrate full stack optimization techniques for robust, reliable AI system performance
PostGIS in Action
PostGIS in Action, Second Edition teaches readers of all levels to write spatial queries that solve real-world problems. It first gives you a background in vector-, raster-, and topology-based GIS and then quickly moves into analyzing, viewing, and mapping data. This second edition covers PostGIS 2.0 and 2.1 series, PostgreSQL 9.1, 9.2, and 9.3 features, and shows you how to integrate with other GIS tools.
Purchase of the print book includes a free eBook in PDF, Kindle, and ePub formats from Manning Publications.
About the Book
Processing data tied to location and topology requires specialized know-how. PostGIS is a free spatial database extender for PostgreSQL, every bit as good as proprietary software. With it, you can easily create location-aware queries in just a few lines of SQL code and build the back end for a mapping, raster analysis, or routing application with minimal effort.
PostGIS in Action, Second Edition teaches you to solve real-world geodata problems. It first gives you a background in vector-, raster-, and topology-based GIS and then quickly moves into analyzing, viewing, and mapping data. You'll learn how to optimize queries for maximum speed, simplify geometries for greater efficiency, and create custom functions for your own applications. You'll also learn how to apply your existing GIS knowledge to PostGIS and integrate with other GIS tools.
Familiarity with relational database and GIS concepts is helpful but not required.
What's Inside
An introduction to spatial databases
Geometry, geography, raster, and topology spatial types, functions, and queries
Applying PostGIS to real-world problems
Extending PostGIS to web and desktop applications
Updated for PostGIS 2.x and PostgreSQL 9.x
About the Authors
Regina Obe and Leo Hsu are database consultants and authors. Regina is a member of the PostGIS core development team and the Project Steering Committee.
Table of Contents
PART 1 INTRODUCTION TO POSTGIS
What is a spatial database?
Spatial data types
Spatial reference system considerations
Working with real data
Using PostGIS on the desktop
Geometry and geography functions
Raster functions
PostGIS TIGER geocoder
Geometry relationships
PART 2 PUTTING POSTGIS TO WORK
Proximity analysis
Geometry and geography processing
Raster processing
Building and using topologies
Organizing spatial data
Query performance tuning
PART 3 USING POSTGIS WITH OTHER TOOLS
Extending PostGIS with pgRouting and procedural languages
Using PostGIS in web applications
66Agentic Design Patterns: A Hands-On Guide to Building Intelligent Systems
This book is a practical resource designed to help developers master the art of building sophisticated AI agents. As artificial intelligence evolves from simple reactive programs to autonomous entities capable of understanding context and making complex decisions, this book provides the essential Design Patterns and proven techniques needed to construct intelligent systems effectively. Each of the 21 Design Patterns represents a fundamental building block for creating agents that can perceive their environment, make informed decisions, and execute actions autonomously.Agentic Design Patterns: A Hands-On Guide to Building Intelligent Systems is structured as a comprehensive hands-on guide, with each chapter dedicated to a single agentic pattern. Within each chapter, you will find a detailed pattern overview, practical applications and use cases, one or more hands-on code example, and key takeaways for quick review. From foundational concepts such as Prompt Chaining and Tool Use to advanced topics like Multi-Agent Collaboration and Self-Correction, readers will gain practical knowledge they can immediately apply. While the chapters build on each other, you can also use the book as a handy reference, jumping to patterns that address your specific challenges.To provide a tangible "canvas" for the code examples, this guide utilizes three prominent agent development frameworks: LangChain and its extension LangGraph, which offer a flexible way to build complex operational sequences; Crew AI, which provides a structured framework for orchestrating multiple agents; and the Google Agent Developer Kit (Google ADK), which offers tools for building, evaluating, and deploying agents. By showcasing examples across these tools, you will gain a broad understanding of how these patterns can be applied in any technical environment.Building effective agentic systems requires more than just a powerful language model; it demands structure and design. Agentic patterns provide reusable, battle-tested solutions to common challenges, much like design patterns in software engineering. They offer a common language that makes an agent's logic clearer, more maintainable, and more robust. By the end of this journey, you will possess both the theoretical understanding and the practical skills to implement these 21 essential patterns, enabling you to build more intelligent, capable, and autonomous systems on your chosen development canvas.
OpenCV Essentials
68Making Things See: 3D Vision with Kinect, Processing, Arduino, and Makerbot
Perfect for hobbyists, makers, artists, and gamers, Making Things See shows you how to build every project with inexpensive off-the-shelf components, including the open source Processing programming language and the Arduino microcontroller. You'll learn basic skills that will enable you to pursue your own creative applications with Kinect.
Create Kinect applications on Mac OS X, Windows, or Linux
Track people with pose detection and skeletonization, and use blob tracking to detect objects
Analyze and manipulate point clouds
Make models for design and fabrication, using 3D scanning technology
Use MakerBot, RepRap, or Shapeways to print 3D objects
Delve into motion tracking for animation and games
Build a simple robot arm that can imitate your arm movements
Discover how skilled artists have used Kinect to build fascinating projects
Augmented Human: How Technology Is Shaping the New Reality
If you're a designer, developer, entrepreneur, student, educator, business leader, artist, or simply curious about AR's possibilities, this insightful guide explains how you can become involved with an exciting, fast-moving technology.
You'll explore how:
Computer vision, machine learning, cameras, sensors, and wearables change the way you see the world
Haptic technology syncs what you see with how something feels
Augmented sound and hearables alter the way you listen to your environment
Digital smell and taste augment the way you share and receive information
New approaches to storytelling immerse and engage users more deeply
Users can augment their bodies with electronic textiles, embedded technology, and brain-controlled interfaces
Human avatars can learn our behaviors and act on our behalf
Digital Image Processing Using MATLAB
Hands-On GPU-Accelerated Computer Vision with OpenCV and CUDA: Effective techniques for processing complex image data in real time using GPUs
An Invitation to 3-D Vision: From Images to Geometric Models
73AI Projects with Raspberry Pi: High-performance artificial intelligence for robotics, security, home automation, and vision (Essentials)
With this essential guide, you'll learn how to create all kinds of AI-powered projects on Raspberry Pi!Combining Raspberry Pi hardware with artificial intelligence gives you new ways to interact with the world: turn your surroundings into a playing field with computer vision models, predict and adjust to changes in environmental conditions, and engage with people in natural ways.You’ll learn how to enhance projects you build with Raspberry Pi computers and microcontrollers, with or without AI accelerator hardware. AI Projects with Raspberry Pi shows you how to run models directly on the CPU and on the Raspberry Pi AI HAT+ and AI HAT+ 2 neural network accelerators. Whether you're using a Raspberry Pi Zero 2 W, Raspberry Pi 4, Raspberry Pi 5, or the ultra-low-cost Raspberry Pi Pico, you'll find something in this guide from Raspberry Pi Press.You'll learn how to:Identify and categorise objects in pictures or videoRecognise and generate speechTranslate from one language to anotherUse generative AI models to create images or textTrain models to make predictions from sensor readings, such as temperature and accelerationThis book starts with a quick overview of the basics of artificial intelligence and machine learning, and then moves on to a variety of hands-on projects you can build yourself. It also includes full-colour illustrations as well as source code you can use in your own Raspberry Pi artificial intelligence adventures!The Raspberry Pi Essentials series offers concise, hands-on learning to the most popular activities for Raspberry Pi's computers and add-on boards. Also available in the series:Conquer the command lineSimple electronics with GPIO ZeroMake games with PythonExperiment with the Sense HAT
Algorithms for Image Processing and Computer Vision
Mastering OpenCV 4 with Python: A practical guide covering topics from image processing, augmented reality to deep learning with OpenCV 4 and Python 3.7
Mastering Computer Vision with TensorFlow 2.x: Build advanced computer vision applications using machine learning and deep learning techniques
Machine Learning for OpenCV: Intelligent image processing with Python
Heart of the Machine: Our Future in a World of Artificial Emotional Intelligence
For Readers of Ray Kurzweil and Michio Kaku, a New Look at the Cutting Edge of Artificial Intelligence
Imagine a robotic stuffed animal that can read and respond to a child’s emotional state, a commercial that can recognize and change based on a customer’s facial expression, or a company that can actually create feelings as though a person were experiencing them naturally. Heart of the Machine explores the next giant step in the relationship between humans and technology: the ability of computers to recognize, respond to, and even replicate emotions. Computers have long been integral to our lives, and their advances continue at an exponential rate. Many believe that artificial intelligence equal or superior to human intelligence will happen in the not-too-distance future; some even think machine consciousness will follow. Futurist Richard Yonck argues that emotion, the first, most basic, and most natural form of communication, is at the heart of how we will soon work with and use computers.
Instilling emotions into computers is the next leap in our centuries-old obsession with creating machines that replicate humans. But for every benefit this progress may bring to our lives, there is a possible pitfall. Emotion recognition could lead to advanced surveillance, and the same technology that can manipulate our feelings could become a method of mass control. And, as shown in movies like Her and Ex Machina, our society already holds a deep-seated anxiety about what might happen if machines could actually feel and break free from our control. Heart of the Machine is an exploration of the new and inevitable ways in which mankind and technology will interact.
Fake Photos (MIT Press Essential Knowledge series)
Stalin, Mao, Hitler, Mussolini, and other dictators routinely doctored photographs so that the images aligned with their messages. They erased people who were there, added people who were not, and manipulated backgrounds. They knew if they changed the visual record, they could change history. Once, altering images required hours in the darkroom; today, it can be done with a keyboard and mouse. Because photographs are so easily faked, fake photos are everywhere--supermarket tabloids, fashion magazines, political ads, and social media. How can we tell if an image is real or false? In this volume in the MIT Press Essential Knowledge series, Hany Farid offers a concise and accessible guide to techniques for detecting doctored and fake images in photographs and digital media.
Farid, an expert in photo forensics, has spent two decades developing techniques for authenticating digital images. These techniques model the entire image-creation process in order to find the digital disruption introduced by manipulation of the image. Each section of the book describes a different technique for analyzing an image, beginning with those requiring minimal technical expertise and advancing to those at intermediate and higher levels. There are techniques for, among other things, reverse image searches, metadata analysis, finding image imperfections introduced by JPEG compression, image cloning, tracing pixel patterns, and detecting images that are computer generated. In each section, Farid describes the techniques, explains when they should be applied, and offers examples of image analysis.
Photo Forensics (The MIT Press)
The first comprehensive and detailed presentation of techniques for authenticating digital images.
Photographs have been doctored since photography was invented. Dictators have erased people from photographs and from history. Politicians have manipulated photos for short-term political gain. Altering photographs in the predigital era required time-consuming darkroom work. Today, powerful and low-cost digital technology makes it relatively easy to alter digital images, and the resulting fakes are difficult to detect. The field of photo forensics—pioneered in Hany Farid's lab at Dartmouth College—restores some trust to photography. In this book, Farid describes techniques that can be used to authenticate photos. He provides the intuition and background as well as the mathematical and algorithmic details needed to understand, implement, and utilize a variety of photo forensic techniques.
Farid traces the entire imaging pipeline. He begins with the physics and geometry of the interaction of light with the physical world, proceeds through the way light passes through a camera lens, the conversion of light to pixel values in the electronic sensor, the packaging of the pixel values into a digital image file, and the pixel-level artifacts introduced by photo-editing software. Modeling the path of light during image creation reveals physical, geometric, and statistical regularities that are disrupted during the creation of a fake. Various forensic techniques exploit these irregularities to detect traces of tampering. A chapter of case studies examines the authenticity of viral video and famously questionable photographs including “Golden Eagle Snatches Kid” and the Lee Harvey Oswald backyard photo.
Superhuman Innovation: Transforming Business with Artificial Intelligence
Artificial Intelligence (AI) is the new electricity of our times. It is revolutionizing industries the world over, and changing how we fundamentally view and understand work. Superhuman Innovation argues that AI will supercharge the workforce and the world of work, can be harnessed to deliver powerful change to how companies innovate and gain competitive advantage. It is a practical guide to how AI and Machine Learning are impacting not only how businesses, brands, and agencies innovate, but also what they innovate: products, services and content.
In a world of product and pricing parity, the delivery of superior service experience has become the new marketing, and the new real competitive edge. With AI companies can harness the power of data, personalization and on-demand availability, at the touch of an intelligent button. Superhuman Innovation discusses how AI will serve the superstar innovators of tomorrow, by enabling them to see deeper insights and set sail for higher goals. It unearths a powerful five-pronged model which describes how AI enables innovation through the offerings of Speed (facilitating work processes), Understanding (revealing and mastering deep insights), Performance (customization of delivery to customers), Experimentation (the iterative process of reinvention and feedback) and Results (tangible, measurable and optimizable results). The book is supported by varied and innovative case studies from a variety of industries.
82Transformers for Natural Language Processing and Computer Vision: Explore Generative AI and Large Language Models with Hugging Face, ChatGPT, GPT-4V, and DALL-E 3
Machine Learning: Make Your Own Recommender System (Machine Learning From Scratch)
Bayesian Reasoning and Machine Learning
85An Illustrated Guide to AI Agents: Concepts and Code for Building Agents with LLMs, Tools, and Memory
Artificial intelligence is entering a new phase. AI agents can now reason, plan, and act with increasing independence. From accelerating scientific breakthroughs to supporting creative work, these systems are quickly reshaping industries and everyday life. This book provides the conceptual foundation and practical insights you need to understand—and effectively work with—this emerging technology.Through hundreds of clear graphic illustrations, Maarten Grootendorst and Jay Alammar explain how AI agents are built, how they think, and where they're heading. Designed for professionals, students, and curious learners alike, this guide goes beyond the buzz to reveal what's actually happening inside these systems, why it matters, and how to apply the knowledge in real-world contexts. An Illustrated Guide to AI Agents is your essential reference for navigating the next frontier of artificial intelligence.Explore the core architecture of AI agents: tools, memory, and planningUnderstand reasoning LLMs, multimodal models, and multi-agent collaborationLearn advanced methods, including distillation, quantization, and reinforcement learningEvaluate real-world applications, strengths, and limitations of AI agents
Hands-On Unsupervised Learning Using Python: How to Build Applied Machine Learning Solutions from Unlabeled Data
Author Ankur Patel shows you how to apply unsupervised learning using two simple, production-ready Python frameworks: Scikit-learn and TensorFlow using Keras. With code and hands-on examples, data scientists will identify difficult-to-find patterns in data and gain deeper business insight, detect anomalies, perform automatic feature engineering and selection, and generate synthetic datasets. All you need is programming and some machine learning experience to get started.
Compare the strengths and weaknesses of the different machine learning approaches: supervised, unsupervised, and reinforcement learning
Set up and manage machine learning projects end-to-end
Build an anomaly detection system to catch credit card fraud
Clusters users into distinct and homogeneous groups
Perform semisupervised learning
Develop movie recommender systems using restricted Boltzmann machines
Generate synthetic images using generative adversarial networks
Explainable AI: Interpreting, Explaining and Visualizing Deep Learning (Lecture Notes in Computer Science)
Fundamentals of Computer Vision
Practical Computer Vision Applications Using Deep Learning with Cnns: With Detailed Examples in Python Using Tensorflow and Kivy
For automating the process, the book highlights the limitations of traditional hand-crafted features for computer vision and why the CNN deep-learning model is the state-of-art solution. CNNs are discussed from scratch to demonstrate how they are different and more efficient than the fully connected ANN (FCNN). You will implement a CNN in Python to give you a full understanding of the model.
After consolidating the basics, you will use TensorFlow to build a practical image-recognition model that you will deploy to a web server using Flask, making it accessible over the Internet. Using Kivy and NumPy, you will create cross-platform data science applications with low overheads.
This book will help you apply deep learning and computer vision concepts from scratch, step-by-step from conception to production.
What You Will Learn
Understand how ANNs and CNNs work
Create computer vision applications and CNNs from scratch using Python
Follow a deep learning project from conception to production using TensorFlow
Use NumPy with Kivy to build cross-platform data science applications
Who This Book Is ForData scientists, machine learning and deep learning engineers, software developers.