<?xml version="1.0" encoding="utf-8" ?><rss version="2.0"><channel><title><![CDATA[weixin_42849849的博客]]></title><description><![CDATA[]]></description><link>https://blog.csdn.net/weixin_42849849</link><language>zh-cn</language><generator>https://blog.csdn.net/</generator><copyright><![CDATA[Copyright &copy; weixin_42849849]]></copyright><item><title><![CDATA[ITHACA‑FV, Reduced order modelling(ROM) techniques for OpenFOAM]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/166786708</link><guid>https://blog.csdn.net/weixin_42849849/article/details/166786708</guid><author>weixin_42849849</author><pubDate>Mon, 28 Sep 2026 14:04:18 +0800</pubDate><description><![CDATA[面向OpenFOAM的开源降阶模型（ROM，Model Order Reduction）C++库，由mathLab课题组开发维护。ITHACA‑FV是构建于OpenFOAM有限体积求解器之上的降阶建模工具库，核心目标：把高保真CFD全阶模型（FOM），通过降阶方法构建代理模型，大幅降低参数化CFD、瞬态流体仿真的计算开销，用于参数扫描、实时预测、数字孪生等场景。传统OpenFOAM直接做多参数算例，每个参数都要跑完整CFD；ITHACA‑FV分为离线阶段。]]></description><category></category></item><item><title><![CDATA[HPC: ARCHER2英国国家超算服务]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/166481290</link><guid>https://blog.csdn.net/weixin_42849849/article/details/166481290</guid><author>weixin_42849849</author><pubDate>Wed, 23 Sep 2026 14:19:44 +0800</pubDate><description><![CDATA[如果你需要，我可以补充：ARCHER2与你现在的EPYC 9654 Zen4硬件关键差异，以及ARCHER2的典型Slurm作业脚本模板。从抓取页面看到的近期活动：2026年9‑10月有多场线上webinar，包括数据迁移、Cirrus新资源过渡、GPU编程、容器技术等。，由UKRI出资，爱丁堡大学EPCC中心运维，硬件为HPE Cray EX，2021年正式上线，接替老一代ARCHER超算。，之后登录、文件系统全部不可访问，用户需要迁移数据、转向Cirrus NCR等后续资源。]]></description><category></category></item><item><title><![CDATA[通信模块 gdrcopy,nccl,nvshemm,ucx,ucc, 及其互相依赖关系]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/166440131</link><guid>https://blog.csdn.net/weixin_42849849/article/details/166440131</guid><author>weixin_42849849</author><pubDate>Wed, 23 Sep 2026 07:50:20 +0800</pubDate><description><![CDATA[推荐组合：OpenMPI + UCX + UCC + NCCL + GDRCopy + NVSHMEM。如果你需要，我可以画Mermaid图，直接嵌入到你的文档里。]]></description><category></category></item><item><title><![CDATA[MPI/HPC高性能计算benchmarks工具资源]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/165489845</link><guid>https://blog.csdn.net/weixin_42849849/article/details/165489845</guid><author>weixin_42849849</author><pubDate>Tue, 15 Sep 2026 20:28:06 +0800</pubDate><description><![CDATA[Intel® MPI Benchmarks
OSU Micro-Benchmarks
NAS Parallel Benchmarks
High Performance Conjugate Gradients (HPCG) Benchmark]]></description><category></category></item><item><title><![CDATA[大型源码仓库构建加速方案]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/164997579</link><guid>https://blog.csdn.net/weixin_42849849/article/details/164997579</guid><author>weixin_42849849</author><pubDate>Fri, 11 Sep 2026 15:39:13 +0800</pubDate><description><![CDATA[Bazel/Buck2 优势：精确文件级依赖追踪、远程分布式缓存、跨机器复用编译产物；同类替代：sccache（Mozilla，支持分布式远程缓存，适合多台构建机器集群，跨主机共享编译产物，适合超大项目）前提：你现在遇到的痛点一般是：全量构建很慢；重大坑：C/C++ 头文件改动，会导致所有 include 这个头的源码重编，这是大头耗时来源。很多大项目慢，不是自己代码，是第三方库（boost、hdf5、mpi等）每次都从头编译。CMake/Ninja 原生支持 PCH，适合C/C++项目。]]></description><category></category></item><item><title><![CDATA[法国EDF电力研究院（EDF R&D）全套开源软件大全]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/162938476</link><guid>https://blog.csdn.net/weixin_42849849/article/details/162938476</guid><author>weixin_42849849</author><pubDate>Thu, 16 Jul 2026 15:18:53 +0800</pubDate><description><![CDATA[SALOME几何建模 → Mesh划分网格 → Code_Aster结构应力计算 / Code_Saturne流体换热 → OpenTURNS不确定性校核 → CloudCompare三维扫描实测模型比对验证。软件，全部面向能源、工业制造、土建、核电场景，大部分开源协议为GPL/LGPL，企业商用免费。EDF作为全球核电、火电、水电龙头，自研/联合开发了一整套。]]></description><category></category></item><item><title><![CDATA[CloudCompare, 3D点云处理软件介绍]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/162938358</link><guid>https://blog.csdn.net/weixin_42849849/article/details/162938358</guid><author>weixin_42849849</author><pubDate>Thu, 16 Jul 2026 15:16:04 +0800</pubDate><description><![CDATA[CloudCompare（简称CC）是完全开源免费、跨平台的三维点云+三角网格处理软件，由法国EDF电力研究院研发，开源协议GPL，GitHub仓库2.1k+星标，是工业逆向、三维检测、激光扫描领域最通用免费工具。官网：https://www.cloudcompare.org支持系统：Windows/macOS/Linux，原生64位，千万级点云流畅加载（底层八叉树加速）核心定位：点云配准、去噪、偏差对比、测量、网格重建、分割。]]></description><category></category></item><item><title><![CDATA[Bojan Niceno博士介绍(CFD/OpenFOAM)]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/162461686</link><guid>https://blog.csdn.net/weixin_42849849/article/details/162461686</guid><author>weixin_42849849</author><pubDate>Tue, 30 Jun 2026 18:06:47 +0800</pubDate><description><![CDATA[https://github.com/Niceno现任职务：教育背景：工作经历：根据他的学术主页和发表记录，其核心研究领域包括：Bojan Niceno博士在TU Delft攻读博士期间，与Muhamed Hadžiabdić共同开发了著名的开源CFD代码T-Flows[[43]][[50]]。该代码是一个二阶精度的有限体积法通用CFD求解器，适用于非结构化网格[[41]][[44]]。目前该代码已在GitHub上开源，采用MIT许可证[[50]][[62]]。他领导了PSI-BOIL等重要项目，专注于：代]]></description><category></category></item><item><title><![CDATA[如何生成动态库的Stub导入库(桩库)]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/162013438</link><guid>https://blog.csdn.net/weixin_42849849/article/details/162013438</guid><author>weixin_42849849</author><pubDate>Mon, 15 Jun 2026 18:59:46 +0800</pubDate><description><![CDATA[Stub桩/导入库生成]]></description><category></category></item><item><title><![CDATA[Python, petsc4py介绍和使用]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161490483</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161490483</guid><author>weixin_42849849</author><pubDate>Thu, 28 May 2026 17:31:31 +0800</pubDate><description><![CDATA[利用 petsc4py，我们用极其简短且优雅的 Python 代码，就撬动了工业界顶级的 MPI+CUDA 混合分布式算力。如果你嫌手写稀疏矩阵太低阶，业界还有完全以 petsc4py 为数学内核的高阶有限元框架，如 FEniCSx 和 Firedrake，非常适合直接拿来做流固耦合（FSI）和共轭传热（CHT）的三维仿真。希望这篇实战硬核教程能帮你避开高性能计算中的那些环境坑，让你的科学计算代码在 GPU 集群上彻底飞起来！]]></description><category></category></item><item><title><![CDATA[Python, CuPy 与 cupyx 入门到实战]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161487344</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161487344</guid><author>weixin_42849849</author><pubDate>Thu, 28 May 2026 15:23:35 +0800</pubDate><description><![CDATA[CuPy 是由 Preferred Networks 开发、基于 NVIDIA CUDA/AMD ROCm 的开源 GPU 数组计算库，最初为深度学习框架 Chainer 后端，如今已是 NumFOCUS 赞助的主流高性能计算项目。核心定位：NumPy 的 GPU 镜像，API 高度兼容，原有 NumPy 代码几乎无需改动即可迁移至 GPU。核心优势：无缝代码迁移、极致运算加速、完整覆盖多维数组、广播、索引、线性代数、FFT、随机数等常用功能。硬件支持。]]></description><category></category></item><item><title><![CDATA[LuaBridge3介绍和使用]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161333388</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161333388</guid><author>weixin_42849849</author><pubDate>Sat, 23 May 2026 07:49:44 +0800</pubDate><description><![CDATA[LuaBridge 3 是目前最轻量、最简单、无依赖的C++ ↔ Lua 绑定库，专门用来在 C++ 里调用 Lua、或者在 Lua 里调用 C++ 函数/类，是游戏、插件系统、脚本化逻辑最常用的库。我给你整理最实用、能直接复制跑的版本，不讲废话。纯头文件库（只需 1 个文件夹，无需编译）支持 C++11 及以上支持 Lua 5.1 / 5.2 / 5.3 / 5.4 / LuaJIT无第三方依赖性能极高、体积极小语法非常简洁，比 LuaBind、Sol2 更轻更快。]]></description><category></category></item><item><title><![CDATA[xtensor-stack 开源组织全解析：背景、核心项目、使用教程]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161333172</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161333172</guid><author>weixin_42849849</author><pubDate>Sat, 23 May 2026 07:33:55 +0800</pubDate><description><![CDATA[是由法国数据分析公司QuantStack主导维护的一套现代 C++ 数值计算生态开源组织，对标 Python NumPy、MATLAB、Eigen，但更侧重现代 C++、向量化 SIMD、零开销抽象、跨平台、高性能科学计算。官网：https://github.com/xtensor-stack把 Python 式简单易用的数组语法，搬到高性能 C++，兼顾易用性 + 极致性能。]]></description><category></category></item><item><title><![CDATA[Illinois Rocstar LLC 完整介绍（CFD/多物理/高性能计算领域）]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161319950</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161319950</guid><author>weixin_42849849</author><pubDate>Fri, 22 May 2026 16:42:42 +0800</pubDate><description><![CDATA[Illinois Rocstar LLC 是美国顶尖的多物理场、高性能并行仿真软件研发与工程服务公司，深耕CFD（计算流体力学）、燃烧、流固耦合、稀疏求解、MPI并行优化，OpenFOAM、稀疏矩阵迭代、幽灵单元交换技术高度契合，是美国能源部、国防部重点合作的仿真技术企业。]]></description><category></category></item><item><title><![CDATA[MPI_Win_allocate_shared介绍和使用]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161319256</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161319256</guid><author>weixin_42849849</author><pubDate>Fri, 22 May 2026 16:14:44 +0800</pubDate><description><![CDATA[MPI_Win_allocate_shared函数]]></description><category></category></item><item><title><![CDATA[单机多卡（多核 CPU + 多张 NVIDIA GPU）环境编写一套现代化 CFD 求解器,技术选型ABC]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161261589</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161261589</guid><author>weixin_42849849</author><pubDate>Wed, 20 May 2026 17:36:32 +0800</pubDate><description><![CDATA[按照这套选型构建的 CFD 求解器，既能在单机多卡环境里压榨出媲美工业商业软件（如 ANSYS Fluent GPU）的极致算力，又能借助 Taskflow 和标准 C++ 的现代化语法，彻底告别过去传统 CFD 代码（如旧版 OpenFOAM 大量的 C 风格指针和宏包裹）那般晦涩难懂的泥潭。CFD 仿真 70% 以上的时间都死卡在压力泊松方程（Poisson Equation）等稀疏线性方程组的迭代求解上。CFD 中的每个网格点都包含坐标（x,y,z）、速度（u,v,w）、压力（p）等大量物理量。]]></description><category></category></item><item><title><![CDATA[Taskflow介绍,核心概念解释]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161259081</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161259081</guid><author>weixin_42849849</author><pubDate>Wed, 20 May 2026 15:49:17 +0800</pubDate><description><![CDATA[你可以把 Taskflow 理解为一个高智能的工程项目经理（Executor）。你只需要拿着画好因果关系的网络图纸（Taskflow）交给它，它就能自动把常规的批量搬砖任务（Parallel For）、突发的紧急子任务（Subflow）、需要看情况决定的审批流（Conditioning）以及外包给 GPU 团队的密集算力任务（cudaFlow）全部井井有条地分配给手下的各个核心，并保证没有任何一个核心在原地死等闲置。]]></description><category></category></item><item><title><![CDATA[vcpkg, 开源的跨平台C/C++包管理器介绍和使用]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161237426</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161237426</guid><author>weixin_42849849</author><pubDate>Wed, 20 May 2026 06:53:45 +0800</pubDate><description><![CDATA[vcpkg介绍]]></description><category></category></item><item><title><![CDATA[浮点数运算里的 Guard Bits（保护位）解释]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161237370</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161237370</guid><author>weixin_42849849</author><pubDate>Wed, 20 May 2026 06:32:51 +0800</pubDate><description><![CDATA[标准尾数：23 bit硬件内部运算尾数：23 + G + R + S =26 bitGuard bits 是浮点数加法单元中，为补偿对阶移位带来的精度丢失，在标准尾数右侧额外添加的若干临时位（G/R/S），用于保存移位后的低位信息，最后按IEEE 754规则舍入，保证运算结果精度符合标准，不产生额外误差。]]></description><category></category></item><item><title><![CDATA[POP: Performance Optimization and Productivity, A Centre of Excellent in HPC]]></title><link>https://blog.csdn.net/weixin_42849849/article/details/161227614</link><guid>https://blog.csdn.net/weixin_42849849/article/details/161227614</guid><author>weixin_42849849</author><pubDate>Tue, 19 May 2026 16:02:17 +0800</pubDate><description><![CDATA[HPC性能优化]]></description><category></category></item></channel></rss>