BLASFEO: basic linear algebra subroutines for embedded optimization
arXiv:1704.02457 · doi:10.1145/3210754
Abstract
BLASFEO is a dense linear algebra library providing high-performance implementations of BLAS- and LAPACK-like routines for use in embedded optimization. A key difference with respect to existing high-performance implementations of BLAS is that the computational performance is optimized for small to medium scale matrices, i.e., for sizes up to a few hundred. BLASFEO comes with three different implementations: a high-performance implementation aiming at providing the highest performance for matrices fitting in cache, a reference implementation providing portability and embeddability and optimized for very small matrices, and a wrapper to standard BLAS and LAPACK providing high-performance on large matrices. The three implementations of BLASFEO together provide high-performance dense linear algebra routines for matrices ranging from very small to large. Compared to both open-source and proprietary highly-tuned BLAS libraries, for matrices of size up to about one hundred the high-performance implementation of BLASFEO is about 20-30% faster than the corresponding level 3 BLAS routines and 2-3 times faster than the corresponding LAPACK routines.
Cited by in corpus (11)
- The Control Toolbox - An Open-Source C++ Library for Robotics, Optimal and Model Predictive Control
- Active Learning of Discrete-Time Dynamics for Uncertainty-Aware Model Predictive Control
- The Linear Algebra Mapping Problem. Current state of linear algebra languages and libraries
- The BLAS API of BLASFEO: optimizing performance for small matrices
- A sparse ADMM-based solver for linear MPC subject to terminal quadratic constraint
- Real-Time Predictive Control for Precision Machining
- Advanced-Step Real-time Iterations with Four Levels -- New Error Bounds and Fast Implementation in acados
- A Machine Learning Approach Towards Runtime Optimisation of Matrix Multiplication
- Performance optimization of BLAS algorithms with band matrices for RISC-V processors
- Time-Optimal Online Replanning for Agile Quadrotor Flight
- Gauss-Newton meets PANOC: A fast and globally convergent algorithm for nonlinear optimal control