ridm@nrct.go.th   ระบบคลังข้อมูลงานวิจัยไทย   รายการโปรดที่คุณเลือกไว้

Optimisation of the enactment of fine-grained distributed data-intensive work flows

หน่วยงาน Edinburgh Research Archive, United Kingdom

รายละเอียด

ชื่อเรื่อง : Optimisation of the enactment of fine-grained distributed data-intensive work flows
นักวิจัย : Liew, Chee Sun
คำค้น : optimisation , enactment , workflow , data-intensive
หน่วยงาน : Edinburgh Research Archive, United Kingdom
ผู้ร่วมงาน : Atkinson, Malcolm , van Hemert, Jano , Cole, Murray , Rodriguez Gonzalez, David , Carpenter, Trevor , Wardlaw, Joanna
ปีพิมพ์ : 2555
อ้างอิง : http://hdl.handle.net/1842/7718
ที่มา : -
ความเชี่ยวชาญ : -
ความสัมพันธ์ : Chee Sun Liew, Amrey Krause, and David Snelling. Dispel enactment, in Malcolm P. Atkinson et al., The DATA Bonanza - Improving Knowledge Discovery for Science, Engineering and Business, Wiley,2012 , Malcolm P. Atkinson, Chee Sun Liew, Michelle Galea, Paul Martin, Amrey Krause, Adrian Mouat, Oscar Corcho, and David Snelling. Data-intensive ar- chitecture for scienti c knowledge discovery, Distributed and Parallel Databases, 30(5), 2012, pp. 307-324. , Chee Sun Liew, Malcolm P. Atkinson, Radoslaw Ostrowski, Murray Cole, Jano I. van Hemert and Liangxiu Han. Performance database: capturing data for op- timizing distributed streaming work ows, Philosophical Transactions of the Royal Society A, 369 (1949), 2011, pp. 3268-3284. , Liangxiu Han, Chee Sun Liew, Malcolm P. Atkinson, and Jano I. van Hemert. A generic parallel processing model for facilitating data mining and integration, Parallel Computing, 37 (3), 2011, pp. 157-171. , Gagarine Yaikhom, Chee Sun Liew, Liangxiu Han, Jano van Hemert, Malcolm Atkinson, and Amy Krause. Federated enactment of work ow patterns, in Euro- Par 2010 - Parallel Processing, P. D'Ambra, M. Guarracino, and D. Talia, eds., vol. 6271 of Lecture Notes in Computer Science, Springer Berlin / Heidelberg, 2010, pp. 317-328. , Chee Sun Liew, Malcolm P. Atkinson, Jano I. van Hemert, and Liangxiu Han. Towards optimising distributed data streaming graphs using parallel streams, in HPDC '10: Proceedings of the 19th ACM International Symposium on High Performance Distributed Computing, S. Hariri and K. Keahey, eds., New York, NY, USA, 2010, ACM, pp. 725-736. , Malcolm P. Atkinson, Jano I. van Hemert, Liangxiu Han, Ally Hume, and Chee Sun Liew. A distributed architecture for data mining and integration, in DADC '09: Proceedings of the second international workshop on Data-aware distributed computing, ACM, 2009, pp. 11-20.
ขอบเขตของเนื้อหา : -
บทคัดย่อ/คำอธิบาย :

The emergence of data-intensive science as the fourth science paradigm has posed a data deluge challenge for enacting scientific work-flows. The scientific community is facing an imminent flood of data from the next generation of experiments and simulations, besides dealing with the heterogeneity and complexity of data, applications and execution environments. New scientific work-flows involve execution on distributed and heterogeneous computing resources across organisational and geographical boundaries, processing gigabytes of live data streams and petabytes of archived and simulation data, in various formats and from multiple sources. Managing the enactment of such work-flows not only requires larger storage space and faster machines, but the capability to support scalability and diversity of the users, applications, data, computing resources and the enactment technologies. We argue that the enactment process can be made efficient using optimisation techniques in an appropriate architecture. This architecture should support the creation of diversified applications and their enactment on diversified execution environments, with a standard interface, i.e. a work-flow language. The work-flow language should be both human readable and suitable for communication between the enactment environments. The data-streaming model central to this architecture provides a scalable approach to large-scale data exploitation. Data-flow between computational elements in the scientific work-flow is implemented as streams. To cope with the exploratory nature of scientific work-flows, the architecture should support fast work-flow prototyping, and the re-use of work-flows and work-flow components. Above all, the enactment process should be easily repeated and automated. In this thesis, we present a candidate data-intensive architecture that includes an intermediate work-flow language, named DISPEL. We create a new fine-grained measurement framework to capture performance-related data during enactments, and design a performance database to organise them systematically. We propose a new enactment strategy to demonstrate that optimisation of data-streaming work-flows can be automated by exploiting performance data gathered during previous enactments.

บรรณานุกรม :
Liew, Chee Sun . (2555). Optimisation of the enactment of fine-grained distributed data-intensive work flows.
    กรุงเทพมหานคร : Edinburgh Research Archive, United Kingdom .
Liew, Chee Sun . 2555. "Optimisation of the enactment of fine-grained distributed data-intensive work flows".
    กรุงเทพมหานคร : Edinburgh Research Archive, United Kingdom .
Liew, Chee Sun . "Optimisation of the enactment of fine-grained distributed data-intensive work flows."
    กรุงเทพมหานคร : Edinburgh Research Archive, United Kingdom , 2555. Print.
Liew, Chee Sun . Optimisation of the enactment of fine-grained distributed data-intensive work flows. กรุงเทพมหานคร : Edinburgh Research Archive, United Kingdom ; 2555.