PromptOptMe: Error-Aware Prompt Compression for LLM-based MT Evaluation Metrics
Evaluating the quality of machine-generated natural language content is a challenging task in Natural Language Processing (NLP). Recently, large language models (LLMs) like GPT-4 have been employed for this purpose, but they are computationally expensive due to the extensive token usage required by…