You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
- This release is mainly focused on improving the multimodal architecture, making MTMD chat handling cleaner, more flexible, and easier to extend for future image, audio, and video workflows. I also continued syncing with the latest llama.cpp APIs and improved several developer-facing diagnostics around model templates, shared library loading, and Windows OpenMP runtime discovery.
Signed-off-by: JamePeng <jame_peng@sina.com>
- refactor(mtmd): extract prompt rendering and media marker normalization
13
+
- Add extra_template_arguments to MTMD chat handlers and pass them through to the Jinja chat template render call. This allows generic model templates to receive render-time options such as enable_thinking, add_vision_id, or model-specific template jinja variables.
14
+
- Extract MTMD prompt rendering into dedicated helpers:
15
+
*_render_mtmd_prompt() for pure chat template rendering
16
+
*_replace_media_placeholders() for normalizing rendered media tags and URLs into the MTMD runtime marker
17
+
*_render_and_replace_media() for the combined render-and-normalize stage
18
+
- This removes inline render/replace logic from _process_mtmd_prompt(), keeps media marker validation after normalization, and improves separation between prompt construction and MTMD tokenization.
- docs(README): Added command prompt scenario for README.md (by **@patrikpatrik**)
98
+
- Updated command prompt scenario under Configuration -> Environment Variables
99
+
- Sanity checking after successful installation of wheel
100
+
101
+
- feat(MTMDChatHandler): add chunk type helpers
102
+
- Add small helper methods `_is_text_chunk`/`_is_image_chunk`/`_is_audio_chunk` for checking
103
+
MTMD text, image, and audio chunk type enum values.
104
+
- This keeps MTMD prompt processing easier to read and avoids repeating direct
105
+
enum comparisons when building token spans for text and media chunks.
106
+
107
+
- feat(mtmd): add video input support to `MTMDChatHandler`
108
+
- Add video_url handling to the MTMD chat template and media extraction
109
+
pipeline. Detect whether the loaded libmtmd build supports video helpers
110
+
and reject video inputs early when MTMD_VIDEO is unavailable.
111
+
- Update media loading and bitmap creation for the new helper wrapper API.
112
+
mtmd_helper_bitmap_init_from_buf now returns a bitmap wrapper containing
113
+
both the decoded bitmap and an optional video helper context, so keep the
114
+
video context alive until mtmd_tokenize completes and release it afterward.
115
+
- Also consolidate duplicated audio/video byte loading into a shared
116
+
_load_bytes helper, reuse it for image loading, and add richer default HTTP
117
+
headers for remote media requests.
118
+
119
+
- build(CMakelists): Improve Windows LLVM OpenMP runtime `libomp140.x86_64.dll` discovery
120
+
- Also improve diagnostics by reporting the selected runtime source and path,
121
+
warning when an explicit override points to a missing file, and keeping a clear
122
+
runtime warning when no OpenMP DLL can be found.
123
+
- prefer VS 2022 VC143 OpenMP redist and keep System32 as final fallback。
124
+
125
+
- feat(_ctypes_extensions): improve error diagnostics for shared library loading
126
+
When `load_shared_library` fails, the resulting `RuntimeError` now
127
+
includes a listing of the contents of the searched directories. This
128
+
provides immediate context to help developers diagnose missing, misplaced,
129
+
or incorrectly named library files.
130
+
131
+
- Added `_format_library_dir_contents` to safely format directory listings.
132
+
- Appended the directory listing to the failure message.
133
+
- Confined this diagnostic work strictly to the failure path to avoid any
134
+
performance overhead during successful imports.
135
+
136
+
- feat: Update llama.cpp to [ggml-org/llama.cpp/commit/3899b39ce2acc2e019f149b7107f24b6ca297390](https://github.com/ggml-org/llama.cpp/commit/3899b39ce2acc2e019f149b7107f24b6ca297390)
137
+
138
+
- feat: Sync llama.cpp llama/mtmd/ggml API Binding 20260707
139
+
140
+
More information see: https://github.com/JamePeng/llama-cpp-python/compare/12861b918f67b62f78f28c5cabb7223f766e1097...b9b58594023ab673c2dda6723f8909d85d65a2e5
0 commit comments