Dipika
AI & ML interests
I work on LLM Compressor, Compressed-Tensors and Speculators
Recent Activity
new activity 1 day ago
RedHatAI/gemma-4-26B-A4B-it-NVFP4:on RTX 3090 new activity 1 day ago
RedHatAI/gemma-4-26B-A4B-it-NVFP4:Impossible to deploy using vLLM 0.20.0 new activity 1 day ago
RedHatAI/gemma-4-26B-A4B-it-NVFP4:Why is max-model-len 96000?