| 文件 | 最后提交记录 | 最后更新时间 |
|---|---|---|
perf: 18x faster read_pixels — convert from cached memory, reuse staging buffer (#397) read_pixels indexed the mapped staging buffer byte-by-byte during BGRA->RGB conversion. Mapped readback memory is uncached (write-combined), so scalar reads run at ~10 MB/s — 99 ms for a 640x480 frame on an RTX 5080, dwarfing the 4 ms render. memcpy each row into a cached local buffer first and convert from there: 5.4 ms (18x). Also: reuse the staging buffer across calls (grown on demand) instead of allocating per call, and wait on the copy's submission index instead of polling the device indefinitely. Per-frame capture loop (sim + render + snap, 640x480): 113 ms -> 12 ms. Co-authored-by: Claude Fable 5 <noreply@anthropic.com> | 2 个月前 | |
feat: switch to wgpu + lines/points rendering + recording feature | 9 个月前 |
| 文件 | 最后提交记录 | 最后更新时间 |
|---|---|---|
| 2 个月前 | ||
| 9 个月前 |