LEADER 02583nam 2200565Ia 450 001 9910462329403321 005 20200520144314.0 010 $a1-283-63622-0 010 $a0-309-25635-6 035 $a(CKB)2670000000241223 035 $a(EBL)3378995 035 $a(SSID)ssj0000665815 035 $a(PQKBManifestationID)11447237 035 $a(PQKBTitleCode)TC0000665815 035 $a(PQKBWorkID)10646852 035 $a(PQKB)10916736 035 $a(MiAaPQ)EBC3378995 035 $a(Au-PeEL)EBL3378995 035 $a(CaPaEBR)ebr10594220 035 $a(CaONFJC)MIL394868 035 $a(OCoLC)798361535 035 $a(EXLCZ)992670000000241223 100 $a20120405d2012 uy 0 101 0 $aeng 135 $aurcn||||||||| 181 $ctxt 182 $cc 183 $acr 200 00$aAssessing the reliability of complex models$b[electronic resource] $emathematical and statistical foundations of verification, validation, and uncertainty quantification /$fCommittee on Mathematical Foundations of Verification, Validation, and Uncertainty Quantification ; Board on Mathematical Sciences and Their Applications ; Division on Engineering and Physical Sciences, National Research Council of the National Academies 210 $aWashington, D.C. $cNational Academies Press$dc2012 215 $a1 online resource (145 p.) 300 $aDescription based upon print version of record. 311 $a0-309-25634-8 320 $aIncludes bibliographical references. 327 $a""Front Matter""; ""Acknowledgments""; ""Contents""; ""Summary""; ""1 Introduction""; ""2 Sources of Uncertainty and Error""; ""3 Verification""; ""4 Emulation, Reduced-Order Modeling, and Forward Propagation""; ""5 Model Validation and Prediction""; ""6 Making Decisions""; ""7 Next Steps in Practice, Research, and Education for Verification, Validation, and Uncertainty Quantification""; ""Appendixes""; ""Appendix A: Glossary""; ""Appendix B: Agendas of Committee Meetings""; ""Appendix C: Committee Biographies""; ""Appendix D: Acronyms"" 606 $aComputer simulation 606 $aUncertainty$xMathematical models 608 $aElectronic books. 615 0$aComputer simulation. 615 0$aUncertainty$xMathematical models. 676 $a003.3 712 02$aNational Research Council (U.S.) 712 02$aNational Academies Press (U.S.) 801 0$bMiAaPQ 801 1$bMiAaPQ 801 2$bMiAaPQ 906 $aBOOK 912 $a9910462329403321 996 $aAssessing the reliability of complex models$91917174 997 $aUNINA LEADER 08338nam 22007815 450 001 996465474703316 005 20200703011855.0 010 $a3-319-74896-3 024 7 $a10.1007/978-3-319-74896-2 035 $a(CKB)4100000002045896 035 $a(DE-He213)978-3-319-74896-2 035 $a(MiAaPQ)EBC6284250 035 $a(MiAaPQ)EBC5592588 035 $a(Au-PeEL)EBL5592588 035 $a(OCoLC)1027063574 035 $a(PPN)224637738 035 $a(EXLCZ)994100000002045896 100 $a20180130d2018 u| 0 101 0 $aeng 135 $aurnn|008mamaa 181 $ctxt$2rdacontent 182 $cc$2rdamedia 183 $acr$2rdacarrier 200 10$aAccelerator Programming Using Directives$b[electronic resource] $e4th International Workshop, WACCPD 2017, Held in Conjunction with the International Conference for High Performance Computing, Networking, Storage and Analysis, SC 2017, Denver, CO, USA, November 13, 2017, Proceedings /$fedited by Sunita Chandrasekaran, Guido Juckeland 205 $a1st ed. 2018. 210 1$aCham :$cSpringer International Publishing :$cImprint: Springer,$d2018. 215 $a1 online resource (IX, 183 p. 59 illus.) 225 1 $aProgramming and Software Engineering ;$v10732 311 $a3-319-74895-5 327 $aIntro -- Preface -- Organization -- Contents -- Applications -- An Example of Porting PETSc Applications to Heterogeneous Platforms with OpenACC -- Abstract -- 1 Introduction -- 2 Workflow and System Description -- 2.1 Workflow -- 2.2 System -- 3 Results and Discussion -- 3.1 Profiling with Score-P -- 3.2 The Most Expensive Kernel: MatMult_SeqAIJ -- 3.3 Four Steps Toward the Final Version of OpenACC Kernel -- 4 Speedups and Strong Scaling -- 5 Conclusion -- Acknowledgement -- References -- Hybrid Fortran: High Productivity GPU Porting Framework Applied to Japanese Weather Prediction Model -- 1 Introduction -- 1.1 ASUCA on GPU -- 1.2 Parallelization Granularity -- 1.3 Memory Layout -- 1.4 Related Work -- 1.5 Problem Summary -- 2 Hybrid Fortran Language Extension and Code Transformation -- 2.1 Parallel Loop Abstraction -- 2.2 Compile-Time Defined Memory Layout and Device Data Region -- 2.3 Transformed Code -- 3 Code Transformation Method -- 4 Productivity- and Performance Results -- 5 Conclusion and Future Work -- References -- Implicit Low-Order Unstructured Finite-Element Multiple Simulation Enhanced by Dense Computation Using OpenACC -- 1 Introduction -- 2 Finite-Element Earthquake Simulation Designed for the K Computer -- 3 Proposed Solver for GPUs Using OpenACC -- 3.1 Modification of Algorithm for GPUs -- 3.2 Introduction of OpenACC -- 4 Performance Measurements -- 5 Application Example -- 6 Concluding Remarks -- References -- Runtime Environments -- The Design and Implementation of OpenMP 4.5 and OpenACC Backends for the RAJA C++ Performance Portability Layer -- 1 Introduction -- 2 RAJA -- 2.1 Basic Execution Policies -- 2.2 RAJA::NestedPolicy and Loop Transformations -- 3 Embedding Directives in the C++ Type System -- 3.1 Defining Policy Tags for a Backend -- 3.2 Constructing Explicit Execution Policy Types. 327 $a3.3 Implement forall Specializations -- 4 Case Study: OpenMP 4.5 -- 5 Case Study: OpenACC -- 6 Evaluation -- 6.1 Test Set -- 6.2 Goals and Non-Goals -- 6.3 Compilation Overhead -- 6.4 Runtime Overhead -- 7 Future Work and Conclusion -- References -- Enabling GPU Support for the COMPSs-Mobile Framework -- 1 Introduction -- 2 Related Work -- 3 Programming Model -- 3.1 Extension for GPU Support -- 4 Runtime Support Implementation -- 4.1 COMPSs-Mobile Runtime Architecture -- 4.2 OpenCL Platform -- 5 Performance Evaluation -- 5.1 OpenCL Platform Performance -- 5.2 Load Balancing Policies -- 6 Conclusions and Future Work -- References -- Concurrent Parallel Processing on Graphics and Multicore Processors with OpenACC and OpenMP -- Abstract -- 1 Introduction -- 2 MBFLO3 Application -- 2.1 Mathematical Formulation -- 2.2 Numerical Method -- 3 Heterogeneous Multiblock Computing Strategy -- 3.1 Multicore Host Parallelism -- 3.2 Manycore Accelerator Parallelism -- 3.3 Heterogeneous Host-Device Parallelism -- 4 Performance Results and Analysis -- 5 Conclusions -- Acknowledgements -- References -- Program Evaluation -- Exploration of Supervised Machine Learning Techniques for Runtime Selection of CPU vs. GPU Execution in Java Programs -- 1 Introduction -- 2 Motivation -- 3 Compiling Java to GPUs -- 3.1 Java Parallel Stream API -- 3.2 JIT Compilation for GPUs -- 4 Exploring Supervised Machine Learning Algorithms -- 4.1 Supervised Machine Learning -- 4.2 Generating Subsets of Features -- 4.3 Constructing Prediction Models -- 4.4 Integrating Prediction Models -- 5 Experimental Results -- 5.1 Experimental Protocol -- 5.2 Overall Summary -- 5.3 Accuracies on the Full Set of Features -- 5.4 Exploring ML Algorithms by Feature Subsetting -- 5.5 Lessons Learned -- 6 Related Work -- 6.1 GPU Code Generation from High-Level Languages -- 6.2 Offline Model Construction. 327 $a7 Conclusions -- A Appendix -- References -- Automatic Testing of OpenACC Applications -- 1 Introduction -- 2 Testing a GPU Port of a Numerical Application -- 3 Autocompare with OpenACC -- 4 Autocompare Implementation -- 5 Experiments -- 6 Related Work -- 7 Future Work -- 8 Conclusion -- References -- Evaluation of Asynchronous Offloading Capabilities of Accelerator Programming Models for Multiple Devices -- 1 Introduction -- 2 Related Work -- 3 Accelerator Programming Models -- 3.1 CUDA -- 3.2 OpenCL -- 3.3 OpenACC -- 3.4 OpenMP -- 4 Implementing the Conjugate Gradient Method -- 5 Performance Results on NVIDIA GPUs -- 5.1 Data Transfers with the Host -- 5.2 Single Device -- 5.3 Two Devices -- 6 Performance Results on Intel Xeon Phi Coprocessors -- 6.1 Single Device -- 6.2 Two Devices -- 7 Summary -- References -- Author Index. 330 $aThis book constitutes the refereed post-conference proceedings of the 4th International Workshop on Accelerator Programming Using Directives, WACCPD 2017, held in Denver, CO, USA, in November 2017. The 9 full papers presented have been carefully reviewed and selected from 14 submissions. The papers share knowledge and experiences to program emerging complex parallel computing systems. They are organized in the following three sections: applications; environments; and program evaluation. 410 0$aProgramming and Software Engineering ;$v10732 606 $aProgramming languages (Electronic computers) 606 $aLogic design 606 $aOperating systems (Computers) 606 $aComputer programming 606 $aComputer organization 606 $aComputers 606 $aProgramming Languages, Compilers, Interpreters$3https://scigraph.springernature.com/ontologies/product-market-codes/I14037 606 $aLogic Design$3https://scigraph.springernature.com/ontologies/product-market-codes/I12050 606 $aOperating Systems$3https://scigraph.springernature.com/ontologies/product-market-codes/I14045 606 $aProgramming Techniques$3https://scigraph.springernature.com/ontologies/product-market-codes/I14010 606 $aComputer Systems Organization and Communication Networks$3https://scigraph.springernature.com/ontologies/product-market-codes/I13006 606 $aModels and Principles$3https://scigraph.springernature.com/ontologies/product-market-codes/I18016 615 0$aProgramming languages (Electronic computers). 615 0$aLogic design. 615 0$aOperating systems (Computers). 615 0$aComputer programming. 615 0$aComputer organization. 615 0$aComputers. 615 14$aProgramming Languages, Compilers, Interpreters. 615 24$aLogic Design. 615 24$aOperating Systems. 615 24$aProgramming Techniques. 615 24$aComputer Systems Organization and Communication Networks. 615 24$aModels and Principles. 676 $a004.3 702 $aChandrasekaran$b Sunita$4edt$4http://id.loc.gov/vocabulary/relators/edt 702 $aJuckeland$b Guido$4edt$4http://id.loc.gov/vocabulary/relators/edt 801 0$bMiAaPQ 801 1$bMiAaPQ 801 2$bMiAaPQ 906 $aBOOK 912 $a996465474703316 996 $aAccelerator Programming Using Directives$91999205 997 $aUNISA