ZICQ
中 Log in / Sign up
Newsroom LLMs & Foundation Models #DASA #LLM Fine-Tuning #ArXiv #Synthetic Data #Model Optimization

ArXiv Research Proposes DASA: Efficient LLM Fine-Tuning Without Human-Readable Text

Avatar of Mr.Xu

By Mr.Xu

Published:

中文阅读 (Chinese) English Version

Summary:ArXiv introduces a novel method for large language model (LLM) fine-tuning called Desired-Update-Aligned Synthetic Data (DASA). This approach leverages activation-gradient feedback from a frozen reference model to optimize continuous synthetic input embeddings, eliminating the need for human-readable text. Experiments on models from the Llama and Qwen families, across six benchmarks, demonstrate that DASA achieves comparable performance to natural language data and surpasses it in several config


Background and Motivation

In recent years, large language models (LLMs) have made significant strides in natural language processing tasks. However, their fine-tuning process typically relies on large amounts of high-quality human-readable text data. This dependency not only increases the difficulty of data acquisition but also introduces potential biases and privacy issues. As a result, researchers are exploring methods for fine-tuning without relying on human-readable text.

Method Overview

This study proposes a novel method called Desired-Update-Aligned Synthetic Data (DASA). DASA utilizes activation-gradient feedback from a frozen reference model to guide the optimization of continuous synthetic input embeddings. The process involves the following steps:

  1. Activation Gradient Feedback: Obtain activation gradients from the frozen reference model to evaluate the adaptability of the synthetic embeddings.
  2. Optimize Synthetic Embeddings: Adjust the synthetic embeddings based on the activation gradients to achieve optimal adaptation for the target task.
  3. Direct Fine-Tuning: Use the optimized embeddings directly for downstream task fine-tuning, with discrete token projections employed only for qualitative inspection.

Experiments and Results

The research team conducted experiments on six models from the Llama and Qwen families, with parameter sizes ranging from 1B to 32B, across six benchmarks spanning knowledge, mathematical reasoning, code generation, and commonsense reasoning. The results demonstrate that:

  • Comparable Performance: DASA achieves performance comparable to natural language data under matched LoRA adaptation settings.
  • Superior Performance: In several configurations, DASA outperforms the existing method GRADMM.
  • Speed Improvement: DASA provides a 3.6 to 4.9 times speedup over GRADMM with comparable peak GPU memory usage.

Technical Highlights

  • No Human-Readable Text Required: DASA eliminates the need for high-quality human-readable text data, reducing data acquisition difficulties.
  • Efficient Optimization: Achieves precise optimization of synthetic embeddings through activation gradient feedback.
  • Wide Applicability: Demonstrates strong performance across multiple benchmarks, showcasing its broad applicability.

Industry Impact and Developer Recommendations

DASA offers a new approach to LLM fine-tuning, particularly in data-scarce or privacy-sensitive application scenarios. Developers can experiment with applying DASA to their own model fine-tuning tasks to improve efficiency and reduce data dependency. Additionally, the speed improvement of DASA opens up new possibilities for real-time application scenarios.


Source: ArXiv AI (cs.AI) (2026-09-30)

— END —

Tags: #DASA #LLM Fine-Tuning #ArXiv #Synthetic Data #Model Optimization

Community Comments

Loading live comments and annotations…