REVIEW 2 cited by
Breeze-7B Technical Report
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Breeze-7B is an open-source language model based on Mistral-7B, designed to address the need for improved language comprehension and chatbot-oriented capabilities in Traditional Chinese. This technical report provides an overview of the additional pretraining, finetuning, and evaluation stages for the Breeze-7B model. The Breeze-7B family of base and chat models exhibits good performance on language comprehension and chatbot-oriented tasks, reaching the top in several benchmarks among models comparable in its complexity class.
Forward citations
Cited by 2 Pith papers
-
Characterizing Bias: Benchmarking Large Language Models in Simplified versus Traditional Chinese
A new benchmark shows LLMs are more accurate in Simplified Chinese for regional terms but favor Taiwanese names in simulated hiring, revealing task-dependent bias between Chinese script variants.
-
Low-Resource Fine-Tuning for Multi-Task Structured Information Extraction with a Billion-Parameter Instruction-Tuned Model
A 1B LLaMA model fine-tuned with LoRA on 100-1000 synthetic samples shows high ROUGE-L and JSON parse rates on three extraction tasks, but the comparison against zero-shot 7B/8B models does not support the claim that ...
Discussion (0). Continue with ORCID to comment.