T3Q-LLM
/

T3Q-LLM2-FP-v1.0

@@ -37,83 +37,19 @@ This is the model card of a 🤗 transformers model that has been pushed on the
 <!-- Address questions around how the model is intended to be used, including the foreseeable users of the model and those affected by the model. -->
-### Direct Use
-<!-- This section is for the model use without fine-tuning or plugging into a larger ecosystem/app. -->
-[More Information Needed]
-### Downstream Use [optional]
-<!-- This section is for the model use when fine-tuned for a task, or when plugged into a larger ecosystem/app -->
-[More Information Needed]
-### Out-of-Scope Use
-<!-- This section addresses misuse, malicious use, and uses that the model will not work well for. -->
-[More Information Needed]
-## Bias, Risks, and Limitations
-<!-- This section is meant to convey both technical and sociotechnical limitations. -->
-[More Information Needed]
-### Recommendations
-<!-- This section is meant to convey recommendations with respect to the bias, risk, and technical limitations. -->
-Users (both direct and downstream) should be made aware of the risks, biases and limitations of the model. More information needed for further recommendations.
-## How to Get Started with the Model
-Use the code below to get started with the model.
-[More Information Needed]
-## Training Details
-### Training Data
-<!-- This should link to a Dataset Card, perhaps with a short stub of information on what the training data is all about as well as documentation related to data pre-processing or additional filtering. -->
-[More Information Needed]
-### Training Procedure
-<!-- This relates heavily to the Technical Specifications. Content here should link to that section when it is relevant to the training procedure. -->
-#### Preprocessing [optional]
-[More Information Needed]
-#### Training Hyperparameters
-- **Training regime:** [More Information Needed] <!--fp32, fp16 mixed precision, bf16 mixed precision, bf16 non-mixed precision, fp16 non-mixed precision, fp8 mixed precision -->
-#### Speeds, Sizes, Times [optional]
-<!-- This section provides information about throughput, start/end time, checkpoint size if relevant, etc. -->
-[More Information Needed]
 ## Evaluation
 <!-- This section describes the evaluation protocols and provides the results. -->
-hf-causal-experimental (pretrained=T3Q-LLM/T3Q-LLM2-FP-v1.0,use_accelerate=true,trust_remote_code=true), limit: None, provide_description: False, num_fewshot: 50, batch_size: 1
 |      Task      |Version| Metric |Value |   |Stderr|
 |----------------|------:|--------|-----:|---|-----:|
-|kobest_boolq    |      0|acc     |0.9466|±  |0.0060|
-|                |       |macro_f1|0.9466|±  |0.0060|
-|kobest_copa     |      0|acc     |0.8680|±  |0.0107|
-|                |       |macro_f1|0.8678|±  |0.0107|
-|kobest_hellaswag|      0|acc     |0.5020|±  |0.0224|
-|                |       |acc_norm|0.5540|±  |0.0223|
-|                |       |macro_f1|0.5007|±  |0.0223|
-|kobest_sentineg |      0|acc     |0.9521|±  |0.0107|
-|                |       |macro_f1|0.9521|±  |0.0108|

 <!-- Address questions around how the model is intended to be used, including the foreseeable users of the model and those affected by the model. -->
 ## Evaluation
 <!-- This section describes the evaluation protocols and provides the results. -->
+hf-causal-experimental (pretrained=T3Q-LLM/T3Q-LLM2-FP-v1.0,use_accelerate=true,trust_remote_code=true), limit: None, provide_description: False, num_fewshot: 0, batch_size: 8
 |      Task      |Version| Metric |Value |   |Stderr|
 |----------------|------:|--------|-----:|---|-----:|
+|kobest_boolq    |      0|acc     |0.5976|±  |0.0131|
+|                |       |macro_f1|0.5224|±  |0.0136|
+|kobest_copa     |      0|acc     |0.8190|±  |0.0122|
+|                |       |macro_f1|0.8189|±  |0.0122|
+|kobest_hellaswag|      0|acc     |0.5240|±  |0.0224|
+|                |       |acc_norm|0.5740|±  |0.0221|
+|                |       |macro_f1|0.5214|±  |0.0224|
+|kobest_sentineg |      0|acc     |0.7809|±  |0.0208|
+|                |       |macro_f1|0.7786|±  |0.0211|