Brian Dellabetta
bdellabe
AI & ML interests
None yet
Recent Activity
updated a model 2 days ago
nm-testing/SmolLM-1.7B-Instruct-quantized.w4a16 updated a model 3 days ago
bdellabe/Inkling-NVFP4-FP8-BLOCK published a model 4 days ago
bdellabe/Inkling-NVFP4-FP8-BLOCKOrganizations
Chat template updates
1
#6 opened 9 days ago
by
Mazyod
Update chat template from upstream model
1
#3 opened 9 days ago
by
lehmanju
Update chat_template.jinja
#4 opened 4 days ago
by
bdellabe
Add mtp regex to ignore list
2
#3 opened 11 days ago
by
bdellabe
Can you add more details to the model card?
6
#1 opened 3 months ago
by
bullerwins
Trouble disabling multi-modality
1
#10 opened about 2 months ago
by
mohamedelyazidferhat
Could you also provide FP8-block version of the gemma-4-31B-it drafted model?
3
#5 opened 3 months ago
by
RayHuang1991
Update config.json
#1 opened about 2 months ago
by
bdellabe
Request: Update Chat Template to match Base Model
3
#6 opened 3 months ago
by
yasu-oh
Update chat_template.jinja
#6 opened 2 months ago
by
bdellabe
Update chat_template.jinja
#8 opened 2 months ago
by
bdellabe
Turboquant and mtp?
2
#8 opened 3 months ago
by
tasticleeze
Update chat_template.jinja
1
#4 opened 3 months ago
by
bdellabe
Update chat_template.jinja
#5 opened 3 months ago
by
bdellabe
fix: add missing use_deterministic_attn parameter to MoonViT3dEncoder
5
#22 opened 3 months ago
by
ace-coreweave
Error KeyError: 'layers.0.experts.0.down_proj.input_global_scale' when running on vllm
1
#5 opened 3 months ago
by
pachePizza
Running Gemma 4 Truthfully at 128K on One RTX 5090
3
#4 opened 3 months ago
by
Mosai-Sys
Problem on Nvidia DGX Spark
1
#4 opened 3 months ago
by
akalongman
Inference is slow with A100
1
#3 opened 3 months ago
by
weisunding
Update README.md
1
#2 opened 3 months ago
by
bullerwins