Tool Matching Rate: An evaluation metric calculated as the number of tools correctly used divided by the total number of tools that should be used.
Fixed-answer: Queries with a static, fill-in-the-blank response.
Open-ended: Queries requiring detailed responses without a predetermined format.
Operational: Queries necessitating the execution of tool-based APIs for specific tasks, evaluated by execution success rate.
Real-time: Queries involving answers that are dynamic over time, evaluated via dynamic assessment.
ReAct: Reasoning and Acting—a paradigm where the agent interleaves reasoning traces with task-specific actions.
Sentence-BERT: A modification of the pre-trained BERT network that uses siamese and triplet network structures to derive semantically meaningful sentence embeddings.
NDCG: Normalized Discounted Cumulative Gain—a metric used to measure the quality of a set of search results.
QLoRA: Quantized Low-Rank Adaptation—a parameter-efficient fine-tuning method that uses quantized weights to reduce memory usage.
LoRA: Low-Rank Adaptation—a technique to fine-tune large models efficiently by updating only a small set of low-rank matrices.
SFT: Supervised Fine-Tuning—training a model on labeled examples to adapt it to specific tasks or instructions.