Syndetics cover image
Image from Syndetics

Problem-solving in high performance computing : a situational awareness approach with Linux / Igor Ljubuncic.

By: Material type: TextTextPublisher: Waltham, Massachusetts : Morgan Kaufmann, 2015Copyright date: ©2015Description: 1 online resource (322 pages) : illustrationsContent type:
  • text
Media type:
  • computer
Carrier type:
  • online resource
ISBN:
  • 9780128010648 (e-book)
Subject(s): Genre/Form: Additional physical formats: Print version:: Problem-solving in high performance computing : a situational awareness approach with Linux.DDC classification:
  • 005.432 23
LOC classification:
  • QA76.76.O63 .L583 2015
Online resources:
Star ratings
    Average rating: 0.0 (0 votes)
Holdings
Item type Current library Call number Status Date due Barcode Item holds
Ebrary Online Books Ebrary Online Books Colombo Available CBERA1000988
Ebrary Online Books Ebrary Online Books Jaffna Available JFEBRA1000988
Ebrary Online Books Ebrary Online Books Kandy Available KDEBRA1000988
Total holds: 0

Enhanced descriptions from Syndetics:

Problem-Solving in High Performance Computing: A Situational Awareness Approach with Linux focuses on understanding giant computing grids as cohesive systems. Unlike other titles on general problem-solving or system administration, this book offers a cohesive approach to complex, layered environments, highlighting the difference between standalone system troubleshooting and complex problem-solving in large, mission critical environments, and addressing the pitfalls of information overload, micro, and macro symptoms, also including methods for managing problems in large computing ecosystems.The authors offer perspective gained from years of developing Intel-based systems that lead the industry in the number of hosts, software tools, and licenses used in chip design. The book offers unique, real-life examples that emphasize the magnitude and operational complexity of high performance computer systems.- Provides insider perspectives on challenges in high performance environments with thousands of servers, millions of cores, distributed data centers, and petabytes of shared data- Covers analysis, troubleshooting, and system optimization, from initial diagnostics to deep dives into kernel crash dumps- Presents macro principles that appeal to a wide range of users and various real-life, complex problems- Includes examples from 24/7 mission-critical environments with specific HPC operational constraints

Includes index.

Description based on online resource; title from PDF title page (EBL, viewed December 2, 2016).

Electronic reproduction. Ann Arbor, MI : ProQuest, 2016. Available via World Wide Web. Access may be limited to ProQuest affiliated libraries.

There are no comments on this title.

to post a comment.