home.social

#fluidx3d — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #fluidx3d, aggregated by home.social.

  1. Directly found the first issue im #FluidX3D: #OpenCL __attribute__((opencl_unroll_hint(x))) is not supported on AMD's old #GPU driver. Guaranteeing legacy hardware compatibility means testing on legacy hardware. I'm not giving software rot a chance! 🖖🧐
    github.com/ProjectPhysX/FluidX

  2. Directly found the first issue im #FluidX3D: #OpenCL __attribute__((opencl_unroll_hint(x))) is not supported on AMD's old #GPU driver. Guaranteeing legacy hardware compatibility means testing on legacy hardware. I'm not giving software rot a chance! 🖖🧐
    github.com/ProjectPhysX/FluidX

  3. Directly found the first issue im #FluidX3D: #OpenCL __attribute__((opencl_unroll_hint(x))) is not supported on AMD's old #GPU driver. Guaranteeing legacy hardware compatibility means testing on legacy hardware. I'm not giving software rot a chance! 🖖🧐
    github.com/ProjectPhysX/FluidX

  4. Directly found the first issue im #FluidX3D: #OpenCL __attribute__((opencl_unroll_hint(x))) is not supported on AMD's old #GPU driver. Guaranteeing legacy hardware compatibility means testing on legacy hardware. I'm not giving software rot a chance! 🖖🧐
    github.com/ProjectPhysX/FluidX

  5. Directly found the first issue im #FluidX3D: #OpenCL __attribute__((opencl_unroll_hint(x))) is not supported on AMD's old #GPU driver. Guaranteeing legacy hardware compatibility means testing on legacy hardware. I'm not giving software rot a chance! 🖖🧐
    github.com/ProjectPhysX/FluidX

  6. #FluidX3D #CFD v3.7 brings faster Q-criterion isosurface rendering with #OpenCL local memory optimization! 🖖🤠
    github.com/ProjectPhysX/FluidX

    Instead of 32 velocities for each #GPU thread, now an 8x8x8 workgroup loads & reuses 11x11x11 velocities in L1$, a 12x VRAM BW reduction.

    Fascinating insight: Which thread loads which cell from VRAM to L1$, and which thread renders which grid cell within the workgroup, can be very different!
    github.com/ProjectPhysX/FluidX

    PS: plugged X-wing Gif in #GitHub preview 🖖😜

  7. #FluidX3D #CFD v3.7 brings faster Q-criterion isosurface rendering with #OpenCL local memory optimization! 🖖🤠
    github.com/ProjectPhysX/FluidX

    Instead of 32 velocities for each #GPU thread, now an 8x8x8 workgroup loads & reuses 11x11x11 velocities in L1$, a 12x VRAM BW reduction.

    Fascinating insight: Which thread loads which cell from VRAM to L1$, and which thread renders which grid cell within the workgroup, can be very different!
    github.com/ProjectPhysX/FluidX

    PS: plugged X-wing Gif in #GitHub preview 🖖😜

  8. #FluidX3D #CFD v3.7 brings faster Q-criterion isosurface rendering with #OpenCL local memory optimization! 🖖🤠
    github.com/ProjectPhysX/FluidX

    Instead of 32 velocities for each #GPU thread, now an 8x8x8 workgroup loads & reuses 11x11x11 velocities in L1$, a 12x VRAM BW reduction.

    Fascinating insight: Which thread loads which cell from VRAM to L1$, and which thread renders which grid cell within the workgroup, can be very different!
    github.com/ProjectPhysX/FluidX

    PS: plugged X-wing Gif in #GitHub preview 🖖😜

  9. #FluidX3D #CFD v3.7 brings faster Q-criterion isosurface rendering with #OpenCL local memory optimization! 🖖🤠
    github.com/ProjectPhysX/FluidX

    Instead of 32 velocities for each #GPU thread, now an 8x8x8 workgroup loads & reuses 11x11x11 velocities in L1$, a 12x VRAM BW reduction.

    Fascinating insight: Which thread loads which cell from VRAM to L1$, and which thread renders which grid cell within the workgroup, can be very different!
    github.com/ProjectPhysX/FluidX

    PS: plugged X-wing Gif in #GitHub preview 🖖😜

  10. #FluidX3D #CFD v3.7 brings faster Q-criterion isosurface rendering with #OpenCL local memory optimization! 🖖🤠
    github.com/ProjectPhysX/FluidX

    Instead of 32 velocities for each #GPU thread, now an 8x8x8 workgroup loads & reuses 11x11x11 velocities in L1$, a 12x VRAM BW reduction.

    Fascinating insight: Which thread loads which cell from VRAM to L1$, and which thread renders which grid cell within the workgroup, can be very different!
    github.com/ProjectPhysX/FluidX

    PS: plugged X-wing Gif in #GitHub preview 🖖😜

  11. Newest #IntelArc #GPU family member is here, the Panther Lake Arc B390... and it... purrs? 🖖 🥺 🐈‍⬛
    My OpenCL-Benchmark on the B390 measures ~7.4 TFlops FP32 and ~120GB/s memory bandwidth. hw-smi also works with the B390.
    #FluidX3D benchmarks here: github.com/ProjectPhysX/FluidX
    And the #OpenCL infos:
    - Arc B390: opencl.gpuinfo.org/displayrepo
    - Core Ultra X7 358H: opencl.gpuinfo.org/displayrepo

  12. Newest #IntelArc #GPU family member is here, the Panther Lake Arc B390... and it... purrs? 🖖 🥺 🐈‍⬛
    My OpenCL-Benchmark on the B390 measures ~7.4 TFlops FP32 and ~120GB/s memory bandwidth. hw-smi also works with the B390.
    #FluidX3D benchmarks here: github.com/ProjectPhysX/FluidX
    And the #OpenCL infos:
    - Arc B390: opencl.gpuinfo.org/displayrepo
    - Core Ultra X7 358H: opencl.gpuinfo.org/displayrepo

  13. Newest #IntelArc #GPU family member is here, the Panther Lake Arc B390... and it... purrs? 🖖 🥺 🐈‍⬛
    My OpenCL-Benchmark on the B390 measures ~7.4 TFlops FP32 and ~120GB/s memory bandwidth. hw-smi also works with the B390.
    #FluidX3D benchmarks here: github.com/ProjectPhysX/FluidX
    And the #OpenCL infos:
    - Arc B390: opencl.gpuinfo.org/displayrepo
    - Core Ultra X7 358H: opencl.gpuinfo.org/displayrepo

  14. Newest #IntelArc #GPU family member is here, the Panther Lake Arc B390... and it... purrs? 🖖 🥺 🐈‍⬛
    My OpenCL-Benchmark on the B390 measures ~7.4 TFlops FP32 and ~120GB/s memory bandwidth. hw-smi also works with the B390.
    #FluidX3D benchmarks here: github.com/ProjectPhysX/FluidX
    And the #OpenCL infos:
    - Arc B390: opencl.gpuinfo.org/displayrepo
    - Core Ultra X7 358H: opencl.gpuinfo.org/displayrepo

  15. Newest #IntelArc #GPU family member is here, the Panther Lake Arc B390... and it... purrs? 🖖 🥺 🐈‍⬛
    My OpenCL-Benchmark on the B390 measures ~7.4 TFlops FP32 and ~120GB/s memory bandwidth. hw-smi also works with the B390.
    #FluidX3D benchmarks here: github.com/ProjectPhysX/FluidX
    And the #OpenCL infos:
    - Arc B390: opencl.gpuinfo.org/displayrepo
    - Core Ultra X7 358H: opencl.gpuinfo.org/displayrepo

  16. #FluidX3D #CFD has reached ⭐ 5000 Stargazers on #GitHub! 🖖🥳
    Grid refinement update is still in development, I haven't forgotten... ⬜◻️◽▫️
    github.com/ProjectPhysX/FluidX

  17. #FluidX3D #CFD has reached ⭐ 5000 Stargazers on #GitHub! 🖖🥳
    Grid refinement update is still in development, I haven't forgotten... ⬜◻️◽▫️
    github.com/ProjectPhysX/FluidX

  18. #FluidX3D #CFD has reached ⭐ 5000 Stargazers on #GitHub! 🖖🥳
    Grid refinement update is still in development, I haven't forgotten... ⬜◻️◽▫️
    github.com/ProjectPhysX/FluidX

  19. #FluidX3D #CFD has reached ⭐ 5000 Stargazers on #GitHub! 🖖🥳
    Grid refinement update is still in development, I haven't forgotten... ⬜◻️◽▫️
    github.com/ProjectPhysX/FluidX

  20. #FluidX3D #CFD has reached ⭐ 5000 Stargazers on #GitHub! 🖖🥳
    Grid refinement update is still in development, I haven't forgotten... ⬜◻️◽▫️
    github.com/ProjectPhysX/FluidX

  21. Finally Intel #GPU support on Linux too. Watch all the metrics go brrr in multi-GPU #FluidX3D #CFD workload! Will #opensource soon™️

    Hardening against the myriads of broken counters in all those bugged APIs was a long shot. 🖖🫠

    ____________ | Windows | #Linux |
    CPU / RAM | ✅️️WinAPI | ✅️️/proc |
    #Nvidia GPU | ✅️️NVML | ✅️️NVML |
    #Intel GPU | ✅IGCL | ✅SYSMAN |
    #AMD GPU | ✅️️️️ADLX | ✅️️️️AMDSMI |

  22. Finally Intel #GPU support on Linux too. Watch all the metrics go brrr in multi-GPU #FluidX3D #CFD workload! Will #opensource soon™️

    Hardening against the myriads of broken counters in all those bugged APIs was a long shot. 🖖🫠

    ____________ | Windows | #Linux |
    CPU / RAM | ✅️️WinAPI | ✅️️/proc |
    #Nvidia GPU | ✅️️NVML | ✅️️NVML |
    #Intel GPU | ✅IGCL | ✅SYSMAN |
    #AMD GPU | ✅️️️️ADLX | ✅️️️️AMDSMI |

  23. Finally Intel #GPU support on Linux too. Watch all the metrics go brrr in multi-GPU #FluidX3D #CFD workload! Will #opensource soon™️

    Hardening against the myriads of broken counters in all those bugged APIs was a long shot. 🖖🫠

    ____________ | Windows | #Linux |
    CPU / RAM | ✅️️WinAPI | ✅️️/proc |
    #Nvidia GPU | ✅️️NVML | ✅️️NVML |
    #Intel GPU | ✅IGCL | ✅SYSMAN |
    #AMD GPU | ✅️️️️ADLX | ✅️️️️AMDSMI |

  24. Finally Intel #GPU support on Linux too. Watch all the metrics go brrr in multi-GPU #FluidX3D #CFD workload! Will #opensource soon™️

    Hardening against the myriads of broken counters in all those bugged APIs was a long shot. 🖖🫠

    ____________ | Windows | #Linux |
    CPU / RAM | ✅️️WinAPI | ✅️️/proc |
    #Nvidia GPU | ✅️️NVML | ✅️️NVML |
    #Intel GPU | ✅IGCL | ✅SYSMAN |
    #AMD GPU | ✅️️️️ADLX | ✅️️️️AMDSMI |

  25. Finally Intel #GPU support on Linux too. Watch all the metrics go brrr in multi-GPU #FluidX3D #CFD workload! Will #opensource soon™️

    Hardening against the myriads of broken counters in all those bugged APIs was a long shot. 🖖🫠

    ____________ | Windows | #Linux |
    CPU / RAM | ✅️️WinAPI | ✅️️/proc |
    #Nvidia GPU | ✅️️NVML | ✅️️NVML |
    #Intel GPU | ✅IGCL | ✅SYSMAN |
    #AMD GPU | ✅️️️️ADLX | ✅️️️️AMDSMI |

  26. #FluidX3D #CFD v3.6 is out! This release accumulates a number of small improvements over the last months. Most notably, better interactive graphics support on #macOS with XQuartz. Have fun! 🖖😎🌊🍏
    github.com/ProjectPhysX/FluidX

  27. #FluidX3D #CFD v3.6 is out! This release accumulates a number of small improvements over the last months. Most notably, better interactive graphics support on #macOS with XQuartz. Have fun! 🖖😎🌊🍏
    github.com/ProjectPhysX/FluidX

  28. #FluidX3D #CFD v3.6 is out! This release accumulates a number of small improvements over the last months. Most notably, better interactive graphics support on #macOS with XQuartz. Have fun! 🖖😎🌊🍏
    github.com/ProjectPhysX/FluidX

  29. #FluidX3D #CFD v3.6 is out! This release accumulates a number of small improvements over the last months. Most notably, better interactive graphics support on #macOS with XQuartz. Have fun! 🖖😎🌊🍏
    github.com/ProjectPhysX/FluidX

  30. #FluidX3D #CFD v3.6 is out! This release accumulates a number of small improvements over the last months. Most notably, better interactive graphics support on #macOS with XQuartz. Have fun! 🖖😎🌊🍏
    github.com/ProjectPhysX/FluidX

  31. Some experimentation with ```mermaid ...``` charts in #GitHub #markdown. Turns out you can hack the formatting on the quadrantChart to turn it into an xy-scatter plot with individual point size/coloring/labeling.🖖🧐
    First plot is datasheet memory bandwidth vs. FP32 TFlops/s, second plot is #FluidX3D performance vs. bandwidth, for lots of #GPU​/​#CPU hardware.

  32. Some experimentation with ```mermaid ...``` charts in #GitHub #markdown. Turns out you can hack the formatting on the quadrantChart to turn it into an xy-scatter plot with individual point size/coloring/labeling.🖖🧐
    First plot is datasheet memory bandwidth vs. FP32 TFlops/s, second plot is #FluidX3D performance vs. bandwidth, for lots of #GPU​/​#CPU hardware.

  33. Some experimentation with ```mermaid ...``` charts in #GitHub #markdown. Turns out you can hack the formatting on the quadrantChart to turn it into an xy-scatter plot with individual point size/coloring/labeling.🖖🧐
    First plot is datasheet memory bandwidth vs. FP32 TFlops/s, second plot is #FluidX3D performance vs. bandwidth, for lots of #GPU​/​#CPU hardware.

  34. Some experimentation with ```mermaid ...``` charts in #GitHub #markdown. Turns out you can hack the formatting on the quadrantChart to turn it into an xy-scatter plot with individual point size/coloring/labeling.🖖🧐
    First plot is datasheet memory bandwidth vs. FP32 TFlops/s, second plot is #FluidX3D performance vs. bandwidth, for lots of #GPU​/​#CPU hardware.

  35. Some experimentation with ```mermaid ...``` charts in #GitHub #markdown. Turns out you can hack the formatting on the quadrantChart to turn it into an xy-scatter plot with individual point size/coloring/labeling.🖖🧐
    First plot is datasheet memory bandwidth vs. FP32 TFlops/s, second plot is #FluidX3D performance vs. bandwidth, for lots of #GPU​/​#CPU hardware.

  36. Here's me demoing #Intel Arc Pro B60 #GPU workstations at #SC25 in St. Louis, runnig SolidWorks and #FluidX3D! 🖖😎
    youtube.com/watch?v=Z8yxiyXTi7I

  37. Here's me demoing #Intel Arc Pro B60 #GPU workstations at #SC25 in St. Louis, runnig SolidWorks and #FluidX3D! 🖖😎
    youtube.com/watch?v=Z8yxiyXTi7I

  38. Here's me demoing #Intel Arc Pro B60 #GPU workstations at #SC25 in St. Louis, runnig SolidWorks and #FluidX3D! 🖖😎
    youtube.com/watch?v=Z8yxiyXTi7I

  39. Here's me demoing #Intel Arc Pro B60 #GPU workstations at #SC25 in St. Louis, runnig SolidWorks and #FluidX3D! 🖖😎
    youtube.com/watch?v=Z8yxiyXTi7I

  40. Here's me demoing #Intel Arc Pro B60 #GPU workstations at #SC25 in St. Louis, runnig SolidWorks and #FluidX3D! 🖖😎
    youtube.com/watch?v=Z8yxiyXTi7I

  41. Making progress on >top secret #FluidX3D update< but still a long way to go 🖖🧐

  42. Making progress on >top secret #FluidX3D update< but still a long way to go 🖖🧐

  43. Making progress on >top secret #FluidX3D update< but still a long way to go 🖖🧐

  44. Making progress on >top secret #FluidX3D update< but still a long way to go 🖖🧐

  45. Making progress on >top secret #FluidX3D update< but still a long way to go 🖖🧐