You need to agree to share your contact information to access this model

This repository is publicly accessible, but you have to accept the conditions to access its files and content.

Log in or Sign Up to review the conditions and access this model content.

palmer-007-preview

This model is an important improvement over our previous model. To get early access to this model you need to join to the CEAMFA community by keeping an active ko-fi subscription. You can get more information on how to ๐Ÿ‘‰access this model by clicking here.

Note

  • Before you activate your subscription, remember that this is a base model not a chat assistant!
  • Temp 0 for best results.
  • A system prompt might help if you still want to use it as a chat model.
  • The model has a known preference to use an unused token. We are still investigating if it is happening due to a mismatch in the tokenizer during the knowledge distillation phase or an issue with gguf, we are working on a fix soon.
  • The issue above can be partly solved by using "<" as stop token as long as your use case does not involve html nor similar tag texts.

Benchmarks

Benchmark What it measures Metric palmer-006 This model
ARC-Easy Basic science and general reasoning acc_norm 44.82% 47.69%
ARC-Challenge Harder scientific reasoning acc_norm 29.27% 30.29%
PIQA Physical commonsense acc_norm 63.60% 63.71%
WinoGrande Pronoun resolution and contextual reasoning accuracy 50.36% 50.83%
HellaSwag Plausible continuation and commonsense acc_norm 38.41% 38.50%
ArithMark-3 Elementary arithmetic and quantitative continuation acc_norm 52.70% 56.20%
BananaMind raw Broad language, reasoning, context and code accuracy 66.00% 66.29%
BananaMind Elo Fixed-item capability rating Elo 1124 1126
TextIntent overall strict Exact task completion across all 1,000 public items strict pass 26.60% 31.70%
TextIntent overall quality Graded semantic and constraint quality quality score 50.43% 59.63%
TextIntent generation strict Exact execution of rewriting, formatting and tool tasks strict pass 0.44% 7.78%
TextIntent generation quality Semantic quality of generated answers quality score 55.63% 72.75%
TextIntent Elo Overall fixed-item benchmark rating Elo 833 880

Details

TextIntent is our internal metric for measured usefulness, behaviour and instruction-following capabilities on small language models. Llama.cpp and Ollama compatible early mid-training checkpoint with significant improvements over palmer-006.

Downloads last month
3
GGUF
Model size
91.1M params
Architecture
falcon-h1
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for CEAMFA/palmer-007-preview

Quantized
(1)
this model