Technigma AI
technigmaai
ยท
AI & ML interests
Exploring open-source LLMs, local AI inference, model serving, quantization, and agentic AI. Particularly interested in benchmarking and optimizing models for real-world tool use, long-context workloads, and efficient inference across NVIDIA DGX Spark, Blackwell GPUs, and other local AI hardware. Experimenting with vLLM, GGUF, NVFP4/FP8, MoE models, AI agents, and practical GenAI infrastructure.
Recent Activity
liked a model 2 days ago
Qwen/Qwen3.8-Flash-Next new activity 13 days ago
unsloth/Qwen3.8-27B-NVFP4:Does not work with vllm 0.27.1 (latest)Organizations
None yet