Jesse Gross
94ab428e3f
ggml: Seperate tensor load from backend creation
...
Currently, when the backend is created, the tensors are loaded at the
same time, which is a slow operation. This separates them to be two
steps:
- Create backend, including enumerating tensors and memory allocation
- Loading tensor data
This allows more flexibility in managing model loading.
2025-05-19 09:54:22 -07:00
..
2025-05-08 11:42:14 -07:00
2024-12-10 12:58:06 -08:00
2024-07-26 14:14:48 -07:00
2025-02-28 16:10:43 -08:00
2025-05-19 09:54:22 -07:00
2025-03-28 11:50:22 -07:00
2024-03-14 20:18:06 -07:00
2024-03-14 20:18:06 -07:00
2025-05-08 11:42:14 -07:00
2025-05-19 09:54:22 -07:00
2024-11-05 14:21:45 -08:00
2024-11-05 14:21:45 -08:00
2024-11-05 14:21:45 -08:00
2024-12-31 18:02:30 -08:00
2025-05-19 09:54:22 -07:00
2025-05-08 11:42:14 -07:00
2025-03-28 11:50:22 -07:00
2025-05-13 17:36:02 -07:00
2025-05-13 17:36:02 -07:00
2025-05-19 09:54:22 -07:00
2025-05-12 15:23:31 -07:00
2025-05-06 11:20:48 -07:00
2024-12-31 18:02:30 -08:00
2025-05-08 11:42:14 -07:00
2024-12-31 18:02:30 -08:00
2025-05-08 13:17:30 -07:00
2025-05-13 17:36:02 -07:00
2025-05-08 11:42:14 -07:00
2025-05-13 17:36:02 -07:00
2024-08-09 12:16:19 -07:00
2024-08-09 12:16:19 -07:00
2025-02-04 19:30:49 -08:00