We’re currently very busy writing down our projects, so you have a better idea what we can do.
As we were too busy handling all the projects and the required growth in the past 5+ years, we did not share much what we did.
By repeated request, we’re now writing these down, so you have an idea of what we are capable of. Please come back later for more stories, or follow us on LinkedIn.
The Fastest Payroll System Of The World
At StreamHPC we do several very different types of projects, but this project has been very, very different. In the first place, it was nowhere close to scientific simulation or media processing. Our client, Intersoft solutions, asked us to speed up thousands of payroll calculations on a GPU. They wanted…
We accelerated the OpenCL backend of pyPaSWAS sequence aligner
Last year we accelerated the OpenCL-code in PaSWAS, which is open source software to do DNA/RNA/protein sequence alignment and trimming. It has users world-wide in universities, research groups and industry. Below you’ll find the benchmark results of our acceleration work. You can also test out yourself, as the code is…
Learn about AMD’s PRNG library we developed: rocRAND – includes benchmarks
When CUDA kept having a dominance over OpenCL, AMD introduced HIP – a programming language that closely resembles CUDA. Now it doesn’t take months to port code to AMD hardware, but more and more CUDA-software converts to HIP without problems. The real large and complex code-bases only take a few…
Demo: cartoonizer on an Altera Arria 10 FPGA
It takes quite some effort to program FPGAs using VHDL or Verilog. Since several years Intel/Altera has OpenCL-drivers, with the goal to reduce this effort. OpenCL-on-FPGAs reduced the required effort to a quarter of the time, while also making it easier to alter the specifications during the project. Exactly the…
Bug fixing the MESA 3D drivers
Most of our projects are around performance optimisation, but we’re cleaning up bugs too. This is because you can only speed up software when certain types of bugs are cleared out. A few months ago, we got a different type of request. If we could solve bugs in MESA 3D…
Caffe and Torch7 ported to AMD GPUs, MXnet WIP
Last week AMD released ports of Caffe, Torch and (work-in-progress) MXnet, so these frameworks now work on AMD GPUs. With the Radeon MI6, MI8 MI25 (25 TFLOPS half precision) to be released soonish, it’s ofcourse simply needed to have software run on these high end GPUs. The ports have been…
We have been awarded the Khronos project to upgrade the OpenCL test suite to 2.2!
Some weeks ago we started with implementing the Compiler Test Suite for OpenCL 2.2. The biggest improvement of OpenCL 2.2 is C++ kernels, which originally was planned for 2.1. SPIRV 1.1 is another big improvement. We are very happy to have a part in making OpenCL better! We find OpenCL C++…
How we sped up a flooding simulation 35 times (from 32-core CPU to multi-GPU)
How water moves through an area given a certain pace of instream, can be fully simulated. We got a request to make such simulation faster, as it took already too much time to do moderate simulations. As the customer wanted to be able to have more details, larger areas and more…
Porting Manchester’s UNIFAC to OpenCL@XeonPhi: 160x speedup
As we cannot use the performance results for most of our commercial projects because they contain sensitive data, we were happy that Dr. David Topping from the University of Manchester was so kind to allow us to share the data for the UNIFAC project. The goal for this project was simple: port…
Loading…
Something went wrong. Please refresh the page and/or try again.