Cross-language reference for native inference SDK integration across C++, Python, and C#, focused on FFI, structure layout, buffer ownership, and resource lifecycle.
-
Updated
Aug 9, 2026 - C#
Cross-language reference for native inference SDK integration across C++, Python, and C#, focused on FFI, structure layout, buffer ownership, and resource lifecycle.
Native llama.cpp orchestration for quantized LLM evaluation, drift analysis, and reproducible model comparison.
Portable native inference stack for Mage-Flow-Turbo using stable-diffusion.cpp/sd-cli, Q8_0 DiT GGUF, Qwen3-VL-4B Q4_K_M and a dedicated VAE, with CPU/CUDA backends, CLI, REST API, model verification and Kaggle integration.
Go library for running the Cactus Needle on-device tool-calling model.
To associate your repository with the native-inference topic, visit your repo's landing page and select "manage topics."