Tag: #onnx
Writing on GPUs, LLMs, MLOps, Kubernetes — and mindset · 2 posts
Edge AI Complete Guide 2025: On-Device Inference, Model Optimization, TensorRT/ONNX/CoreML
Everything about Edge AI! On-device inference (TensorRT/ONNX Runtime/CoreML/TFLite), model optimization (quantization/pruning/knowledge distillation), hardware (NVIDIA Jetson/Apple Neural Engine/Qualcomm NPU), Federated
2026-04-13 · 21 min read #edge-ai#on-device#inference#tensorrt#onnxEdge AI and On-Device ML Complete Guide: TFLite, ONNX, Core ML, llama.cpp
A complete guide to running AI models on edge devices. Learn hands-on how to deploy optimized AI in mobile and edge environments using TensorFlow Lite, ONNX Runtime, Core ML, MediaPipe, llama.cpp, and Whisper.cpp.
2026-03-17 · 25 min read #edge-ai#on-device-ml#tflite#onnx#core-ml