Technigma AI
technigmaai
ยท
AI & ML interests
Exploring open-source LLMs, local AI inference, model serving, quantization, and agentic AI. Particularly interested in benchmarking and optimizing models for real-world tool use, long-context workloads, and efficient inference across NVIDIA DGX Spark, Blackwell GPUs, and other local AI hardware. Experimenting with vLLM, GGUF, NVFP4/FP8, MoE models, AI agents, and practical GenAI infrastructure.
Recent Activity
liked a model 1 day ago
Qwen/Qwen3.8-Flash-Next new activity 12 days ago
unsloth/Qwen3.8-27B-NVFP4:Does not work with vllm 0.27.1 (latest)Organizations
None yet