Ding-Yong Hong scite author profile

Ding-Yong Hong

4Publications

34Citation Statements Received

29Citation Statements Given

How they've been cited

119

How they cite others

Affiliations

Institute of Information Science, Academia Sinica, Academia Sinica, National Tsing Hua University

Publications

Order By: Most citations

Improving SIMD Parallelism via Dynamic Binary Translation

Hong

Liu

et al. 2018

ACM Trans. Embed. Comput. Syst.

View full text Add to dashboard Cite

Recent trends in SIMD architecture have tended toward longer vector lengths, and more enhanced SIMD features have been introduced in newer vector instruction sets. However, legacy or proprietary applications compiled with short-SIMD ISA cannot benefit from the long-SIMD architecture that supports improved parallelism and enhanced vector primitives, resulting in only a small fraction of potential peak performance. This article presents a dynamic binary translation technique that enables short-SIMD binaries to exploit benefits of new SIMD architectures by rewriting short-SIMD loop code. We propose a general approach that translates loops consisting of short-SIMD instructions to machine-independent IR, conducts SIMD loop transformation/optimization at this IR level, and finally translates to long-SIMD instructions. Two solutions are presented to enforce SIMD load/store alignment, one for the problem caused by the binary translator’s internal translation condition and one general approach using dynamic loop peeling optimization. Benchmark results show that average speedups of 1.51× and 2.48× are achieved for an ARM NEON to x86 AVX2 and x86 AVX-512 loop transformation, respectively.

show abstract

Optimizing Control Transfer and Memory Virtualization in Full System Emulators

Hong

Hsu

Chou

et al. 2015

ACM Trans. Archit. Code Optim.

View full text Add to dashboard Cite

Full system emulators provide virtual platforms for several important applications, such as kernel and system software development, co-verification with cycle accurate CPU simulators, or application development for hardware still in development. Full system emulators usually use dynamic binary translation to obtain reasonable performance. This paper focuses on optimizing the performance of full system emulators. First, we optimize performance by enabling classic control transfer optimizations of dynamic binary translation in full system emulation, such as indirect branch target caching and block chaining. Second, we improve the performance of memory virtualization of cross-ISA virtual machines by improving the efficiency of the software translation lookaside buffer (software TLB). We implement our optimizations on QEMU, an industrial-strength full system emulator, along with the Android emulator. Experimental results show that our optimizations achieve an average speedup of 1.98X for ARM-to-X86-64 QEMU running SPEC CINT2006 benchmarks with train inputs. Our optimizations also achieve an average speedup of 1.44X and 1.40X for IA32-to-X86-64 QEMU and AArch64-to-X86-64 QEMU on SPEC CINT2006. We use a set of real applications downloaded from Google Play as benchmarks for the Android emulator. Experimental results show that our optimizations achieve an average speedup of 1.43X for the Android emulator running these applications. CCS Concepts: r General and reference → Metrics; r Software and its engineering → Virtual machines; Runtime environments

show abstract

Early experiences in application level I/O tracing on blue gene systems

Seelam

Chung

Hong

et al. 2008

View full text Add to dashboard Cite

Hqemu

Hong

Hsu

Yew

et al. 2012

View full text Add to dashboard Cite

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

hi@scite.ai

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Ding-Yong Hong

Improving SIMD Parallelism via Dynamic Binary Translation

Optimizing Control Transfer and Memory Virtualization in Full System Emulators

Early experiences in application level I/O tracing on blue gene systems

Hqemu

Contact Info

Product

Resources

About