TueCICL at SemEval-2024 Task 8: Resource-efficient approaches for machine-generated text detection
Daniel Stuhlinger, Aron Winkler · 2024
Recent developments in the field of NLP have brought large language models (LLMs) to the forefront of both public and research attention.As the use of language generation technologies becomes more widespread, the problem arises of determining whether a given text is machine generated or not.Task 8 at SemEval 2024 consists of a shared task with this exact objective.Our approach aims at developing models and strategies that strike a good balance between performance and model size.We show that it is possible to compete with large transformerbased solutions with smaller systems.Our code can be found on GitHub. 1