9798296089038 - ai engineering: building multi-modal intelligent systems with vision, language, and audio from llm fine-tuning to voice agents, ar interfaces, and real-world deployment di ara, husn (9 risultati)

- Brossura
Da: GreatBookPrices, Columbia, MD, U.S.A.GreatBookPrices
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 31,31
EUR 2,27 spedizioneSpedito in U.S.A.Quantità: Più di 20 disponibili
Condizione: New.

- Brossura
Da: PBShop.store US, Wood Dale, IL, U.S.A.PBShop.store US
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 37,52
Spedizione gratuitaSpedito in U.S.A.Quantità: Più di 20 disponibili
PAP. Condizione: New. New Book. Shipped from UK. Established seller since 2000.

- Brossura
Da: GreatBookPrices, Columbia, MD, U.S.A.GreatBookPrices
Contatta il venditoreVenditore con 5 stelleCondizione: Usato - Come nuovo
EUR 35,70
EUR 2,27 spedizioneSpedito in U.S.A.Quantità: Più di 20 disponibili
Condizione: As New. Unread book in perfect condition.

- Brossura
Da: PBShop.store UK, Fairford, GLOS, Regno UnitoPBShop.store UK
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 34,89
EUR 4,85 spedizioneSpedito da Regno Unito a U.S.A.Quantità: Più di 20 disponibili
PAP. Condizione: New. New Book. Shipped from UK. Established seller since 2000.

- Brossura
Da: GreatBookPricesUK, Woodford Green, Regno UnitoGreatBookPricesUK
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 34,88
EUR 17,48 spedizioneSpedito da Regno Unito a U.S.A.Quantità: Più di 20 disponibili
Condizione: New.

- Brossura
Da: GreatBookPricesUK, Woodford Green, Regno UnitoGreatBookPricesUK
Contatta il venditoreVenditore con 5 stelleCondizione: Usato - Come nuovo
EUR 37,45
EUR 17,48 spedizioneSpedito da Regno Unito a U.S.A.Quantità: Più di 20 disponibili
Condizione: As New. Unread book in perfect condition.

- Brossura
- Print on Demand
Da: California Books, Miami, FL, U.S.A.California Books
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 33,66
Spedizione gratuitaSpedito in U.S.A.Quantità: Più di 20 disponibili
Condizione: New. Print on Demand.

- Brossura
- Print on Demand
Da: Grand Eagle Retail, Bensenville, IL, U.S.A.Grand Eagle Retail
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 37,72
Spedizione gratuitaSpedito in U.S.A.Quantità: 1 disponibili
Paperback. Condizione: new. Paperback. AI Engineering: Building Multi-Modal Intelligent Systems with Vision, Language, and AudioFrom LLM Fine-Tuning to Voice Agents, AR Interfaces, and Real-World DeploymentUnlock the future of artificial intelligence with practical, production-ready multi-modal engineering.This hands-on guide is… built for developers, researchers, and AI professionals who want to go beyond chatbots and dive into building intelligent systems that understand text, images, audio, and human intent - all in one pipeline.Whether you're fine-tuning large language models (LLMs) or creating voice-driven AR interfaces, this book walks you through the real engineering decisions, tools, and architectures needed to bring multi-modal AI to life.What You'll Learn: Fine-tuning Large Language Models (LLMs): Train and adapt models like GPT-2, LLaMA, and Mistral for custom tasks using Hugging Face, LoRA, QLoRA, and PEFT.Voice Interfaces: Combine Whisper, LLMs, and Bark/Tortoise TTS to build interactive speech-driven assistants.Computer Vision + Language: Use models like BLIP, CLIP, and DETR to connect what systems see to what they say and understand.Instruction Tuning & Hyperparameter Optimization: Build smarter, domain-specific models with efficient training workflows.Multi-Modal Pipelines: Chain audio, image, and text inputs for question answering, summarization, tutoring, and AR/robotic control.Real-Time Interfaces: Deploy intelligent agents using FastAPI, Streamlit, Gradio, Docker, and Hugging Face Spaces.Edge & Offline Deployment: Optimize models with ONNX, quantization (4-bit, 8-bit), and TensorRT for low-latency inference on CPU/GPU.Use Cases Covered: Smart document summarizers with OCR + TTSVoice-enabled image assistantsEmotion-aware agentsVirtual tutorsAR-enhanced AI interfacesRobotic perception + control from voice/image inputSecure, multilingual, and privacy-conscious AI systemsTools & Frameworks Inside: Python, PyTorch, Hugging Face TransformersLangChain, OpenCV, Whisper, TTS, BLIPROS, Unity (AR/VR), Gradio, StreamlitDocker, FastAPI, gRPC, TorchServeBuilt for engineers. Written with depth. Designed for real-world impact.If you're ready to build intelligent multi-modal agents that understand the world like humans do - across speech, vision, and language - this book gives you the complete roadmap.Perfect for: Machine learning engineers, data scientists, AI product developers, researchers, robotics engineers, and anyone building cutting-edge AI systems. This item is printed on demand. Shipping may be from multiple locations in the US or from the UK, depending on stock availability.

- Brossura
- Print on Demand
Da: CitiRetail, Stevenage, Regno UnitoCitiRetail
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 38,98
EUR 43,11 spedizioneSpedito da Regno Unito a U.S.A.Quantità: 1 disponibili
Paperback. Condizione: new. Paperback. AI Engineering: Building Multi-Modal Intelligent Systems with Vision, Language, and AudioFrom LLM Fine-Tuning to Voice Agents, AR Interfaces, and Real-World DeploymentUnlock the future of artificial intelligence with practical, production-ready multi-modal engineering.This hands-on guide is… built for developers, researchers, and AI professionals who want to go beyond chatbots and dive into building intelligent systems that understand text, images, audio, and human intent - all in one pipeline.Whether you're fine-tuning large language models (LLMs) or creating voice-driven AR interfaces, this book walks you through the real engineering decisions, tools, and architectures needed to bring multi-modal AI to life.What You'll Learn: Fine-tuning Large Language Models (LLMs): Train and adapt models like GPT-2, LLaMA, and Mistral for custom tasks using Hugging Face, LoRA, QLoRA, and PEFT.Voice Interfaces: Combine Whisper, LLMs, and Bark/Tortoise TTS to build interactive speech-driven assistants.Computer Vision + Language: Use models like BLIP, CLIP, and DETR to connect what systems see to what they say and understand.Instruction Tuning & Hyperparameter Optimization: Build smarter, domain-specific models with efficient training workflows.Multi-Modal Pipelines: Chain audio, image, and text inputs for question answering, summarization, tutoring, and AR/robotic control.Real-Time Interfaces: Deploy intelligent agents using FastAPI, Streamlit, Gradio, Docker, and Hugging Face Spaces.Edge & Offline Deployment: Optimize models with ONNX, quantization (4-bit, 8-bit), and TensorRT for low-latency inference on CPU/GPU.Use Cases Covered: Smart document summarizers with OCR + TTSVoice-enabled image assistantsEmotion-aware agentsVirtual tutorsAR-enhanced AI interfacesRobotic perception + control from voice/image inputSecure, multilingual, and privacy-conscious AI systemsTools & Frameworks Inside: Python, PyTorch, Hugging Face TransformersLangChain, OpenCV, Whisper, TTS, BLIPROS, Unity (AR/VR), Gradio, StreamlitDocker, FastAPI, gRPC, TorchServeBuilt for engineers. Written with depth. Designed for real-world impact.If you're ready to build intelligent multi-modal agents that understand the world like humans do - across speech, vision, and language - this book gives you the complete roadmap.Perfect for: Machine learning engineers, data scientists, AI product developers, researchers, robotics engineers, and anyone building cutting-edge AI systems. This item is printed on demand. Shipping may be from our UK warehouse or from our Australian or US warehouses, depending on stock availability.