• Fri frakt över 249 kr
  • •
  • Snabba leveranser
  • •
  • Billiga böcker
Kundservice

Du är på sajten för privatpersoner.

Företag, bibliotek eller offentlig verksamhet?

Du handlar på classic.bokus.com, där alla dina funktioner finns intakta.
Till classic.bokus.com
Bokus logotyp. Gå till startsidan.
  • Erbjudanden
  • Nyheter
  • Student
  • Topplistor
  • Barn & ungdom
  • Bokus Play
  • E-böcker
  • Pocketböcker
  • Spel & pussel

10% rabatt på allt med kod NYSTART10 →

Sidfot

Mina sidor

    Hjälp

    • Kundservice
    • Vanliga frågor och svar
    • Frakt och leverans
    • Retur vid ångerrätt
    • Reklamera vara
    • Betalning
    • Köpvillkor
    • Allmänna villkor
    • Information om webbplatsens tillgänglighet

    Om Bokus

    • Om oss
    • Pressrum
    • För studenter
    • För företag
    • För bibliotek och offentlig verksamhet
    • För leverantörer
    • Hållbarhet

    Populärt

    • Aktuella erbjudanden
    • Presentkort
    • Studentlitteratur
    • Nya böcker
    • Topplistor
    • Signerade böcker
    • Engelska böcker

    Inspiration

    • Boktips
    • BookTok
    • Populära bokserier
    • Barnbokskaraktärer
    • Populära författare
    Logotyp för Bokus
    Följ oss på Facebook (extern länk)Följ oss på Instagram (extern länk)Följ oss på YouTube (extern länk)Följ oss på TikTok (extern länk)
    bokus @ CookiesAnpassa cookiesIntegritetspolicyKöpvillkor
    Till Citymail hemsida (extern länk)Till Budbee hemsida (extern länk)Till Postnord hemsida (extern länk)Till Schenker hemsida (extern länk)Till Early Bird hemsida (extern länk)Till Walleys hemsida (extern länk)
    1. Data och IT
    2. Programmeringsböcker
    3. Programspråk

    Domain-Specific Computer Architectures for Emerging Applications II

    From Deep Learning to Large Language Models

    AvChao Wang,Wenqi Lou

    Inbunden, Engelska, 2026

    1 265 kr

    Slutsåld

    Beskrivning

    Domain-Specific Computer Architectures for Emerging Applications II: From Deep Learning to Large Language Models provides a systematic account of the architectural shift from task-specific deep learning accelerators to the computing platforms required by transformer-based large language models, Vision Transformers, and mixture-of-experts networks.As AI models scale, performance is determined not only by arithmetic throughput but also by memory bandwidth, data movement, communication efficiency, compiler support, and full-stack hardware-software co-design. This book explains why acceleration strategies developed for convolutional neural networks are no longer sufficient for many contemporary workloads, and presents the architectural principles required for self-attention, sparse execution, KV-cache management, heterogeneous acceleration, and cluster-scale inference. By connecting algorithmic structure with accelerator design, compiler automation, and distributed systems, it offers a unified technical framework for modern AI computing. Topics covered include GPUs, FPGAs, ASICs, spatial accelerators, sparse tensor compilation, auto-tuning, high-level synthesis, neural architecture search, and scalable large language model serving.Combining conceptual foundations with concrete system methodologies, the book is intended for graduate students, researchers, and practitioners in computer architecture, AI systems, and hardware-software co-design seeking to understand how specialized computing platforms are evolving for the foundation-model era.

    Produktinformation

    • Utgivningsdatum:2026-12-24
    • Mått:156 x 234 x undefined mm
    • Format:Inbunden
    • Språk:Engelska
    • Antal sidor:400
    • Förlag:Taylor & Francis Ltd
    • ISBN:9781032952192

    Utforska kategorier

    • Programspråk inom Data och IT
    • Programvaruutveckling inom Data och IT
    • Systemvetenskap och AI inom Data och IT

    Mer om författaren

    Chao Wang is a Professor, Vice Dean of the School of Software, and doctoral supervisor at the University of Science and Technology of China (USTC). His research interests include FPGA-based reconfigurable computing, intelligent processors, and intelligent computing systems. He has led or participated in national and provincial-level research projects and contributed to intelligent computing systems based on domestic AI chips.Wenqi Lou is an Associate Researcher and master's supervisor in the School of Software at USTC. He received his PhD in computer architecture from USTC in 2023. His research interests include intelligent accelerator architectures, FPGA accelerator design, and hardware-software co-optimization for the deployment of deep learning models.Teng Wang is an Associate Researcher in the School of Software at USTC. He received his PhD in computer architecture from USTC in 2023. His research focuses on reconfigurable hardware accelerators and neural network processors.Lei Gong is an Associate Professor in the School of Computer Science and Technology at USTC. He received his PhD in computer science from USTC in 2019. His research interests include FPGA-based accelerator design and artificial intelligence and machine learning systems.Xuehai Zhou is a Professor and doctoral supervisor at USTC. His research interests include heterogeneous multicore architectures, reconfigurable systems, application-specific hardware acceleration, and embedded system design. He has led or participated in more than 30 national research projects.

    Innehållsförteckning

    • 1. Foundations of Domain-Specific Computer Architectures in the Era of Foundation Models 2. An Overview of Large Language Models and Fundamental Algorithms 3. Modeling and Evaluation Framework for Constrained Dataflow in Spatial Accelerators 4. Heterogeneous Acceleration and Adaptive Mapping for Vision Transformers 5. Collaborative Design of MoE ViT: Quantization and Computation Orchestration 6. Scalable Inference Acceleration System for LLM 7. A Compiler Framework for Automatic Generation of High-Performance FPGA Accelerators for Sparse Tensor Computations 8. Tensor Program Auto-Tuning for Tensorized Intrinsics and Diverse GPU 9. High-Level Synthesis Toolchain Enhancement 10. Neural Architecture Search with Zero/Few-Shot Performance Proxies 11. CNN-FPGA Co-Optimization: Synergistic Design of Operator Search, Model Compression, and Hardware Acceleration 12. CNN-ASIC Co-Search and Joint Optimization: Zero-Cost Proxy-Driven Heterogeneous Multi-Core Accelerator Design