---
_id: '63890'
abstract:
- lang: eng
  text: The computation of highly contracted electron repulsion integrals (ERIs) is
    essential to achieve quantum accuracy in atomistic simulations based on quantum
    mechanics. Its growing computational demands make energy efficiency a critical
    concern. Recent studies demonstrate FPGAs’ superior performance and energy efficiency
    for computing primitive ERIs, but the computation of highly contracted ERIs introduces
    significant algorithmic complexity and new design challenges for FPGA acceleration.In
    this work, we present SORCERI, the first streaming overlay acceleration for highly
    contracted ERI computations on FPGAs. SORCERI introduces a novel streaming Rys
    computing unit to calculate roots and weights of Rys polynomials on-chip, and
    a streaming contraction unit for the contraction of primitive ERIs. This shifts
    the design bottleneck from limited CPU-FPGA communication bandwidth to available
    FPGA computation resources. To address practical deployment challenges for a large
    number of quartet classes, we design three streaming overlays, together with an
    efficient memory transpose optimization, to cover the 21 most commonly used quartet
    classes in realistic atomistic simulations. To address the new computation constraints,
    we use flexible calculation stages with a free-running streaming architecture
    to achieve high DSP utilization and good timing closure.Experiments demonstrate
    that SORCERI achieves an average 5.96x, 1.99x, and 1.16x better performance per
    watt than libint on a 64-core AMD EPYC 7713 CPU, libintx on an Nvidia A40 GPU,
    and SERI, the prior best-performing FPGA design for primitive ERIs. Furthermore,
    SORCERI reaches a peak throughput of 44.11 GERIS (109 ERIs per second) that is
    1.52x, 1.13x, and 1.93x greater than libint, libintx and SERI, respectively. SORCERI
    will be released soon at https://github.com/SFU-HiAccel/SORCERI.
author:
- first_name: Philip
  full_name: Stachura, Philip
  last_name: Stachura
- first_name: Xin
  full_name: Wu, Xin
  id: '77439'
  last_name: Wu
- first_name: Christian
  full_name: Plessl, Christian
  id: '16153'
  last_name: Plessl
  orcid: 0000-0001-5728-9982
- first_name: Zhenman
  full_name: Fang, Zhenman
  last_name: Fang
citation:
  ama: 'Stachura P, Wu X, Plessl C, Fang Z. SORCERI: Streaming Overlay Acceleration
    for Highly Contracted Electron Repulsion Integral Computations in Quantum Chemistry.
    In: <i>Proceedings of the 2026 ACM/SIGDA International Symposium on Field Programmable
    Gate Arrays (FPGA ’26)</i>. Association for Computing Machinery; 2026:224-234.
    doi:<a href="https://doi.org/10.1145/3748173.3779198">10.1145/3748173.3779198</a>'
  apa: 'Stachura, P., Wu, X., Plessl, C., &#38; Fang, Z. (2026). SORCERI: Streaming
    Overlay Acceleration for Highly Contracted Electron Repulsion Integral Computations
    in Quantum Chemistry. <i>Proceedings of the 2026 ACM/SIGDA International Symposium
    on Field Programmable Gate Arrays (FPGA ’26)</i>, 224–234. <a href="https://doi.org/10.1145/3748173.3779198">https://doi.org/10.1145/3748173.3779198</a>'
  bibtex: '@inproceedings{Stachura_Wu_Plessl_Fang_2026, place={New York, NY, USA},
    title={SORCERI: Streaming Overlay Acceleration for Highly Contracted Electron
    Repulsion Integral Computations in Quantum Chemistry}, DOI={<a href="https://doi.org/10.1145/3748173.3779198">10.1145/3748173.3779198</a>},
    booktitle={Proceedings of the 2026 ACM/SIGDA International Symposium on Field
    Programmable Gate Arrays (FPGA ’26)}, publisher={Association for Computing Machinery},
    author={Stachura, Philip and Wu, Xin and Plessl, Christian and Fang, Zhenman},
    year={2026}, pages={224–234} }'
  chicago: 'Stachura, Philip, Xin Wu, Christian Plessl, and Zhenman Fang. “SORCERI:
    Streaming Overlay Acceleration for Highly Contracted Electron Repulsion Integral
    Computations in Quantum Chemistry.” In <i>Proceedings of the 2026 ACM/SIGDA International
    Symposium on Field Programmable Gate Arrays (FPGA ’26)</i>, 224–34. New York,
    NY, USA: Association for Computing Machinery, 2026. <a href="https://doi.org/10.1145/3748173.3779198">https://doi.org/10.1145/3748173.3779198</a>.'
  ieee: 'P. Stachura, X. Wu, C. Plessl, and Z. Fang, “SORCERI: Streaming Overlay Acceleration
    for Highly Contracted Electron Repulsion Integral Computations in Quantum Chemistry,”
    in <i>Proceedings of the 2026 ACM/SIGDA International Symposium on Field Programmable
    Gate Arrays (FPGA ’26)</i>, 2026, pp. 224–234, doi: <a href="https://doi.org/10.1145/3748173.3779198">10.1145/3748173.3779198</a>.'
  mla: 'Stachura, Philip, et al. “SORCERI: Streaming Overlay Acceleration for Highly
    Contracted Electron Repulsion Integral Computations in Quantum Chemistry.” <i>Proceedings
    of the 2026 ACM/SIGDA International Symposium on Field Programmable Gate Arrays
    (FPGA ’26)</i>, Association for Computing Machinery, 2026, pp. 224–34, doi:<a
    href="https://doi.org/10.1145/3748173.3779198">10.1145/3748173.3779198</a>.'
  short: 'P. Stachura, X. Wu, C. Plessl, Z. Fang, in: Proceedings of the 2026 ACM/SIGDA
    International Symposium on Field Programmable Gate Arrays (FPGA ’26), Association
    for Computing Machinery, New York, NY, USA, 2026, pp. 224–234.'
date_created: 2026-02-06T06:43:22Z
date_updated: 2026-02-09T09:16:32Z
department:
- _id: '27'
- _id: '518'
doi: 10.1145/3748173.3779198
keyword:
- electron repulsion integrals
- quantum chemistry
- atomistic simulation
- overlay architecture
- fpga acceleration
language:
- iso: eng
main_file_link:
- url: https://dl.acm.org/doi/10.1145/3748173.3779198
page: 224-234
place: New York, NY, USA
project:
- _id: '52'
  name: Computing Resources Provided by the Paderborn Center for Parallel Computing
publication: Proceedings of the 2026 ACM/SIGDA International Symposium on Field Programmable
  Gate Arrays (FPGA '26)
publication_identifier:
  isbn:
  - '9798400720796'
publication_status: published
publisher: Association for Computing Machinery
status: public
title: 'SORCERI: Streaming Overlay Acceleration for Highly Contracted Electron Repulsion
  Integral Computations in Quantum Chemistry'
type: conference
user_id: '77439'
year: '2026'
...
---
_id: '11950'
abstract:
- lang: eng
  text: Advances in electromyographic (EMG) sensor technology and machine learning
    algorithms have led to an increased research effort into high density EMG-based
    pattern recognition methods for prosthesis control. With the goal set on an autonomous
    multi-movement prosthesis capable of performing training and classification of
    an amputee’s EMG signals, the focus of this paper lies in the acceleration of
    the embedded signal processing chain. We present two Xilinx Zynq-based architectures
    for accelerating two inherently different high density EMG-based control algorithms.
    The first hardware accelerated design achieves speed-ups of up to 4.8 over the
    software-only solution, allowing for a processing delay lower than the sample
    period of 1 ms. The second system achieved a speed-up of 5.5 over the software-only
    version and operates at a still satisfactory low processing delay of up to 15
    ms while providing a higher reliability and robustness against electrode shift
    and noisy channels.
author:
- first_name: Alexander
  full_name: Boschmann, Alexander
  last_name: Boschmann
- first_name: Andreas
  full_name: Agne, Andreas
  last_name: Agne
- first_name: Georg
  full_name: Thombansen, Georg
  last_name: Thombansen
- first_name: Linus Matthias
  full_name: Witschen, Linus Matthias
  id: '49051'
  last_name: Witschen
- first_name: Florian
  full_name: Kraus, Florian
  last_name: Kraus
- first_name: Marco
  full_name: Platzner, Marco
  id: '398'
  last_name: Platzner
citation:
  ama: Boschmann A, Agne A, Thombansen G, Witschen LM, Kraus F, Platzner M. Zynq-based
    acceleration of robust high density myoelectric signal processing. <i>Journal
    of Parallel and Distributed Computing</i>. 2019;123:77-89. doi:<a href="https://doi.org/10.1016/j.jpdc.2018.07.004">10.1016/j.jpdc.2018.07.004</a>
  apa: Boschmann, A., Agne, A., Thombansen, G., Witschen, L. M., Kraus, F., &#38;
    Platzner, M. (2019). Zynq-based acceleration of robust high density myoelectric
    signal processing. <i>Journal of Parallel and Distributed Computing</i>, <i>123</i>,
    77–89. <a href="https://doi.org/10.1016/j.jpdc.2018.07.004">https://doi.org/10.1016/j.jpdc.2018.07.004</a>
  bibtex: '@article{Boschmann_Agne_Thombansen_Witschen_Kraus_Platzner_2019, title={Zynq-based
    acceleration of robust high density myoelectric signal processing}, volume={123},
    DOI={<a href="https://doi.org/10.1016/j.jpdc.2018.07.004">10.1016/j.jpdc.2018.07.004</a>},
    journal={Journal of Parallel and Distributed Computing}, publisher={Elsevier},
    author={Boschmann, Alexander and Agne, Andreas and Thombansen, Georg and Witschen,
    Linus Matthias and Kraus, Florian and Platzner, Marco}, year={2019}, pages={77–89}
    }'
  chicago: 'Boschmann, Alexander, Andreas Agne, Georg Thombansen, Linus Matthias Witschen,
    Florian Kraus, and Marco Platzner. “Zynq-Based Acceleration of Robust High Density
    Myoelectric Signal Processing.” <i>Journal of Parallel and Distributed Computing</i>
    123 (2019): 77–89. <a href="https://doi.org/10.1016/j.jpdc.2018.07.004">https://doi.org/10.1016/j.jpdc.2018.07.004</a>.'
  ieee: A. Boschmann, A. Agne, G. Thombansen, L. M. Witschen, F. Kraus, and M. Platzner,
    “Zynq-based acceleration of robust high density myoelectric signal processing,”
    <i>Journal of Parallel and Distributed Computing</i>, vol. 123, pp. 77–89, 2019.
  mla: Boschmann, Alexander, et al. “Zynq-Based Acceleration of Robust High Density
    Myoelectric Signal Processing.” <i>Journal of Parallel and Distributed Computing</i>,
    vol. 123, Elsevier, 2019, pp. 77–89, doi:<a href="https://doi.org/10.1016/j.jpdc.2018.07.004">10.1016/j.jpdc.2018.07.004</a>.
  short: A. Boschmann, A. Agne, G. Thombansen, L.M. Witschen, F. Kraus, M. Platzner,
    Journal of Parallel and Distributed Computing 123 (2019) 77–89.
date_created: 2019-07-12T13:13:55Z
date_updated: 2022-01-06T06:51:13Z
department:
- _id: '78'
doi: 10.1016/j.jpdc.2018.07.004
intvolume: '       123'
keyword:
- High density electromyography
- FPGA acceleration
- Medical signal processing
- Pattern recognition
- Prosthetics
language:
- iso: eng
page: 77-89
publication: Journal of Parallel and Distributed Computing
publication_identifier:
  issn:
  - 0743-7315
publication_status: published
publisher: Elsevier
status: public
title: Zynq-based acceleration of robust high density myoelectric signal processing
type: journal_article
user_id: '398'
volume: 123
year: '2019'
...
---
_id: '9784'
abstract:
- lang: eng
  text: Piezoelectric inertia motors use the inertia of a body to drive it by means
    of a friction contact in a series of small steps. These motors can operate in
    ``stick-slip'' or ``slip-slip'' mode, with the fundamental frequency of the driving
    signal ranging from several Hertz to more than 100 kHz. To predict the motor characteristics,
    a Coulomb friction model is sufficient in many cases, but numerical simulation
    requires microscopic time steps. This contribution proposes a much faster simulation
    technique using one evaluation per period of the excitation signal. The proposed
    technique produces results very close to those of timestep simulation for ultrasonics
    inertia motors and allows direct determination of the steady-state velocity of
    an inertia motor from the motion profile of the driving part. Thus it is a useful
    simulation technique which can be applied in both analysis and design of inertia
    motors, especially for parameter studies and optimisation.
author:
- first_name: Matthias
  full_name: Hunstig, Matthias
  last_name: Hunstig
- first_name: Tobias
  full_name: Hemsel, Tobias
  last_name: Hemsel
- first_name: Walter
  full_name: Sextro, Walter
  last_name: Sextro
citation:
  ama: 'Hunstig M, Hemsel T, Sextro W. An efficient simulation technique for high-frequency
    piezoelectric inertia motors. In: <i>Ultrasonics Symposium (IUS), 2012 IEEE International</i>.
    ; 2012:277-280. doi:<a href="https://doi.org/10.1109/ULTSYM.2012.0068">10.1109/ULTSYM.2012.0068</a>'
  apa: Hunstig, M., Hemsel, T., &#38; Sextro, W. (2012). An efficient simulation technique
    for high-frequency piezoelectric inertia motors. In <i>Ultrasonics Symposium (IUS),
    2012 IEEE International</i> (pp. 277–280). <a href="https://doi.org/10.1109/ULTSYM.2012.0068">https://doi.org/10.1109/ULTSYM.2012.0068</a>
  bibtex: '@inproceedings{Hunstig_Hemsel_Sextro_2012, title={An efficient simulation
    technique for high-frequency piezoelectric inertia motors}, DOI={<a href="https://doi.org/10.1109/ULTSYM.2012.0068">10.1109/ULTSYM.2012.0068</a>},
    booktitle={Ultrasonics Symposium (IUS), 2012 IEEE International}, author={Hunstig,
    Matthias and Hemsel, Tobias and Sextro, Walter}, year={2012}, pages={277–280}
    }'
  chicago: Hunstig, Matthias, Tobias Hemsel, and Walter Sextro. “An Efficient Simulation
    Technique for High-Frequency Piezoelectric Inertia Motors.” In <i>Ultrasonics
    Symposium (IUS), 2012 IEEE International</i>, 277–80, 2012. <a href="https://doi.org/10.1109/ULTSYM.2012.0068">https://doi.org/10.1109/ULTSYM.2012.0068</a>.
  ieee: M. Hunstig, T. Hemsel, and W. Sextro, “An efficient simulation technique for
    high-frequency piezoelectric inertia motors,” in <i>Ultrasonics Symposium (IUS),
    2012 IEEE International</i>, 2012, pp. 277–280.
  mla: Hunstig, Matthias, et al. “An Efficient Simulation Technique for High-Frequency
    Piezoelectric Inertia Motors.” <i>Ultrasonics Symposium (IUS), 2012 IEEE International</i>,
    2012, pp. 277–80, doi:<a href="https://doi.org/10.1109/ULTSYM.2012.0068">10.1109/ULTSYM.2012.0068</a>.
  short: 'M. Hunstig, T. Hemsel, W. Sextro, in: Ultrasonics Symposium (IUS), 2012
    IEEE International, 2012, pp. 277–280.'
date_created: 2019-05-13T13:20:17Z
date_updated: 2022-01-06T07:04:20Z
department:
- _id: '151'
doi: 10.1109/ULTSYM.2012.0068
keyword:
- friction
- ultrasonic motors
- Coulomb friction model
- efficient simulation technique
- friction contact
- high-frequency piezoelectric inertia motor
- motor characteristics prediction
- numerical simulation
- slip-slip mode
- stick-slip mode
- time-step simulation
- ultrasonic inertia motor
- Acceleration
- Acoustics
- Actuators
- Computational modeling
- Friction
- Numerical models
- Steady-state
language:
- iso: eng
page: 277-280
publication: Ultrasonics Symposium (IUS), 2012 IEEE International
publication_identifier:
  issn:
  - 1948-5719
quality_controlled: '1'
status: public
title: An efficient simulation technique for high-frequency piezoelectric inertia
  motors
type: conference
user_id: '55222'
year: '2012'
...
---
_id: '37067'
abstract:
- lang: eng
  text: IP-XACT is a well accepted standard for the exchange of IP components at Electronic
    System and Register Transfer Level. Still, the creation and manipulation of these
    descriptions at the XML level can be time-consuming and error-prone. In this paper,
    we show that the UML can be consistently applied as an efficient and comprehensible
    frontend for IP-XACT-based IP description and integration. For this, we present
    an IP-XACT UML profile that enables UML-based descriptions covering the same information
    as a corresponding IP-XACT description. This enables the automated generation
    of IP-XACT component and design descriptions from respective UML models. In particular,
    it also allows the integration of existing IPs with UML. To illustrate our approach,
    we present an application example based on the IBM PowerPC Evaluation Kit.
author:
- first_name: Tim
  full_name: Schattkowsky, Tim
  last_name: Schattkowsky
- first_name: Tao
  full_name: Xie, Tao
  last_name: Xie
- first_name: Wolfgang
  full_name: Müller, Wolfgang
  id: '16243'
  last_name: Müller
citation:
  ama: 'Schattkowsky T, Xie T, Müller W. A UML Frontend for IP-XACT-based IP Management.
    In: <i>Proceedings of DATE’09</i>. IEEE; 2009. doi:<a href="https://doi.org/10.1109/DATE.2009.5090664">10.1109/DATE.2009.5090664</a>'
  apa: Schattkowsky, T., Xie, T., &#38; Müller, W. (2009). A UML Frontend for IP-XACT-based
    IP Management. <i>Proceedings of DATE’09</i>. Design, Automation &#38; Test in
    Europe Conference &#38; Exhibition. <a href="https://doi.org/10.1109/DATE.2009.5090664">https://doi.org/10.1109/DATE.2009.5090664</a>
  bibtex: '@inproceedings{Schattkowsky_Xie_Müller_2009, place={Nice, France}, title={A
    UML Frontend for IP-XACT-based IP Management}, DOI={<a href="https://doi.org/10.1109/DATE.2009.5090664">10.1109/DATE.2009.5090664</a>},
    booktitle={Proceedings of DATE’09}, publisher={IEEE}, author={Schattkowsky, Tim
    and Xie, Tao and Müller, Wolfgang}, year={2009} }'
  chicago: 'Schattkowsky, Tim, Tao Xie, and Wolfgang Müller. “A UML Frontend for IP-XACT-Based
    IP Management.” In <i>Proceedings of DATE’09</i>. Nice, France: IEEE, 2009. <a
    href="https://doi.org/10.1109/DATE.2009.5090664">https://doi.org/10.1109/DATE.2009.5090664</a>.'
  ieee: 'T. Schattkowsky, T. Xie, and W. Müller, “A UML Frontend for IP-XACT-based
    IP Management,” presented at the Design, Automation &#38; Test in Europe Conference
    &#38; Exhibition, 2009, doi: <a href="https://doi.org/10.1109/DATE.2009.5090664">10.1109/DATE.2009.5090664</a>.'
  mla: Schattkowsky, Tim, et al. “A UML Frontend for IP-XACT-Based IP Management.”
    <i>Proceedings of DATE’09</i>, IEEE, 2009, doi:<a href="https://doi.org/10.1109/DATE.2009.5090664">10.1109/DATE.2009.5090664</a>.
  short: 'T. Schattkowsky, T. Xie, W. Müller, in: Proceedings of DATE’09, IEEE, Nice,
    France, 2009.'
conference:
  name: Design, Automation & Test in Europe Conference & Exhibition
date_created: 2023-01-17T11:54:02Z
date_updated: 2023-01-17T11:54:07Z
department:
- _id: '672'
doi: 10.1109/DATE.2009.5090664
keyword:
- Unified modeling language
- XML
- Power system modeling
- Application software
- Master-slave
- Power system management
- Acceleration
- Scattering
- Software engineering
- Software standards
language:
- iso: eng
place: Nice, France
publication: Proceedings of DATE'09
publication_identifier:
  isbn:
  - 978-1-4244-3781-8
publisher: IEEE
status: public
title: A UML Frontend for IP-XACT-based IP Management
type: conference
user_id: '5786'
year: '2009'
...
---
_id: '2420'
abstract:
- lang: eng
  text: ' This paper presents the acceleration of minimum-cost covering problems by
    instance-specific hardware. First, we formulate the minimum-cost covering problem
    and discuss a branch \& bound algorithm to solve it. Then we describe instance-specific
    hardware architectures that implement branch \& bound in 3-valued logic and use
    reduction techniques similar to those found in software solvers. We further present
    prototypical accelerator implementations and a corresponding design tool flow.
    Our experiments reveal significant raw speedups up to five orders of magnitude
    for a set of smaller unate covering problems. Provided that hardware compilation
    times can be reduced, we conclude that instance-specific acceleration of hard
    minimum-cost covering problems will lead to substantial overall speedups. '
author:
- first_name: Christian
  full_name: Plessl, Christian
  id: '16153'
  last_name: Plessl
  orcid: 0000-0001-5728-9982
- first_name: Marco
  full_name: Platzner, Marco
  id: '398'
  last_name: Platzner
citation:
  ama: Plessl C, Platzner M. Instance-Specific Accelerators for Minimum Covering.
    <i>Journal of Supercomputing</i>. 2003;26(2):109-129. doi:<a href="https://doi.org/10.1023/a:1024443416592">10.1023/a:1024443416592</a>
  apa: Plessl, C., &#38; Platzner, M. (2003). Instance-Specific Accelerators for Minimum
    Covering. <i>Journal of Supercomputing</i>, <i>26</i>(2), 109–129. <a href="https://doi.org/10.1023/a:1024443416592">https://doi.org/10.1023/a:1024443416592</a>
  bibtex: '@article{Plessl_Platzner_2003, title={Instance-Specific Accelerators for
    Minimum Covering}, volume={26}, DOI={<a href="https://doi.org/10.1023/a:1024443416592">10.1023/a:1024443416592</a>},
    number={2}, journal={Journal of Supercomputing}, publisher={Kluwer Academic Publishers},
    author={Plessl, Christian and Platzner, Marco}, year={2003}, pages={109–129} }'
  chicago: 'Plessl, Christian, and Marco Platzner. “Instance-Specific Accelerators
    for Minimum Covering.” <i>Journal of Supercomputing</i> 26, no. 2 (2003): 109–29.
    <a href="https://doi.org/10.1023/a:1024443416592">https://doi.org/10.1023/a:1024443416592</a>.'
  ieee: C. Plessl and M. Platzner, “Instance-Specific Accelerators for Minimum Covering,”
    <i>Journal of Supercomputing</i>, vol. 26, no. 2, pp. 109–129, 2003.
  mla: Plessl, Christian, and Marco Platzner. “Instance-Specific Accelerators for
    Minimum Covering.” <i>Journal of Supercomputing</i>, vol. 26, no. 2, Kluwer Academic
    Publishers, 2003, pp. 109–29, doi:<a href="https://doi.org/10.1023/a:1024443416592">10.1023/a:1024443416592</a>.
  short: C. Plessl, M. Platzner, Journal of Supercomputing 26 (2003) 109–129.
date_created: 2018-04-17T15:10:00Z
date_updated: 2022-01-06T06:56:10Z
department:
- _id: '518'
- _id: '78'
doi: 10.1023/a:1024443416592
extern: '1'
intvolume: '        26'
issue: '2'
keyword:
- reconfigurable computing
- instance-specific acceleration
- minimum covering
language:
- iso: eng
page: 109-129
publication: Journal of Supercomputing
publication_identifier:
  issn:
  - 0920-8542
publisher: Kluwer Academic Publishers
status: public
title: Instance-Specific Accelerators for Minimum Covering
type: journal_article
user_id: '398'
volume: 26
year: '2003'
...
