arxiv.org/abs/1809.03374v1
It is argued that the $x-y$ cancellation model (XYCM) is a good proxy for discussions of finetuned cancellations in physical theories. XYCM is then analyzed from a statistical perspective, where it is argued that a finetuned point in the parameter sp...
www.reddit.com/r/ResearchML/comments/1oyfk2z/i_finetuned_debertav3large_on_mtebamazon_polarity/
TLDR I finetuned 'deberta-v3-large' on 'mteb/amazon\_polarity', got some biased results, for some countries like Iran, Cuba, N. Korea, etc. the results were negative, and for US, EU, India and Russia...
arxiv.org/abs/2512.14751v1
Finetuning pretrained large language models (LLMs) has become the standard paradigm for developing downstream applications. However, its security implications remain unclear, particularly regarding whether finetuned LLMs inherit jailbreak vulnerabili...
github.com/sebastianbk/finetuned-resnet50-keras
This repo shows how to finetune a ResNet50 model for your own data using Keras. (⭐ 163)
arxiv.org/abs/2306.01181v3
Transfer learning has become an increasingly popular technique in machine learning as a way to leverage a pretrained model trained for one task to assist with building a finetuned model for a related task. This paradigm has been especially popular fo...
github.com/alexandrainst/alexandra_ai_eval
Evaluation of finetuned models. (⭐ 9)
arxiv.org/abs/2308.03051v2
Despite the purported multilingual proficiency of instruction-finetuned large language models (LLMs) such as ChatGPT and Bard, the linguistic inclusivity of these models remains insufficiently explored. Considering this constraint, we present a thoro...
arxiv.org/abs/2506.06928v1
Research into Video Large Language Models (LLMs) has progressed rapidly, with numerous models and benchmarks emerging in just a few years. Typically, these models are initialized with a pretrained text-only LLM and finetuned on both image- and video-...
github.com/s-omranpour/Shirin-Sokhan
A Persian Poet Transformer! (finetuned GPT2 on Ganjoor data) (⭐ 5)
arxiv.org/abs/2212.10560v2
Large "instruction-tuned" language models (i.e., finetuned to respond to instructions) have demonstrated a remarkable ability to generalize zero-shot to new tasks. Nevertheless, they depend heavily on human-written instruction data that is often limi...
arxiv.org/abs/2403.05562v1
A transformative approach to mental health therapy lies at the crossroads of cultural heritage and advanced technology. This paper introduces an innovative method that fuses machine learning techniques with traditional Emirati motifs, focusing on the...
arxiv.org/abs/2307.08067v1
Supersymmetric models with low electroweak finetuning are expected to be more prevalent on the string landscape than finetuned models. We assume a fertile patch of landscape vacua containing the minimal supersymmetric standard model (MSSM) as low ene...
arxiv.org/abs/2506.20746v3
When an LLM learns a new fact during finetuning (e.g., new movie releases, newly elected pope, etc.), where does this information go? Are entities enriched with relation information immediately, or do models recall information just-in-time before a p...
github.com/lukasgarbas/nlp-text-emotion
Multi-class sentiment analysis lstm, finetuned bert (⭐ 221)
github.com/punica-ai/punica
Serving multiple LoRA finetuned LLM as one (⭐ 1144)
arxiv.org/abs/2410.19889v1
Finetuning is a common practice widespread across different communities to adapt pretrained models to particular tasks. Text classification is one of these tasks for which many pretrained models are available. On the other hand, ensembles of neural n...
github.com/xlang-ai/instructor-embedding
[ACL 2023] One Embedder, Any Task: Instruction-Finetuned Text Embeddings (⭐ 2023)
www.reddit.com/r/singularity/comments/1iqw3w6/grok_3_was_finetuned_as_a_right_wing_propaganda/
...
arxiv.org/abs/2402.03660v2
The pretraining-finetuning paradigm has become the prevailing trend in modern deep learning. In this work, we discover an intriguing linear phenomenon in models that are initialized from a common pretrained checkpoint and finetuned on different tasks...
arxiv.org/abs/2209.13146v2
The studies of predicting affective states from human voices have relied heavily on speech. This study, indeed, explores the recognition of humans' affective state from their vocal burst, a short non-verbal vocalization. Borrowing the idea from the r...