Fetching the paper…
Reading the bibliography…
In this work, we present a comprehensive exploration of finetuning Malaysian language models, specifically Llama2 and Mistral, on embedding tasks involving negative and positive pairs.
Learning a similarity metric discriminatively, with application to face verification
S. Chopra, R. Hadsell, and Y. LeCun · 2005
Earlier work this paper cites.
Billion-scale similarity search with gpus, 2017
Jeff Johnson, Matthijs Douze, and Hervé Jégou · 2017
Earlier work this paper cites.
C-pack: Packaged resources to advance general chinese embedding, 2023
Shitao Xiao, Zheng Liu, Peitian Zhang, and Niklas Muennighoff · 2023
Cited alongside, same era.
Large malaysian language model based on mistral for enhanced local language understanding, 2024
Husein Zolkepli, Aisyah Razak, Kamarul Adha, and Ariff Nazhan · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…