---
title: Dataset Size Recovery from LoRA Weights
url: https://www.ml-quant.com/papers/arxiv/2406.19395/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2406.19395
source_url: https://arxiv.org/abs/2406.19395
featured: 2024-07-03
citations: 9
topic: Other
---


# Dataset Size Recovery from LoRA Weights

A new task called dataset size recovery aims to determine the number of samples used to train a model, with a method called DSiRe proposed for this purpose.

- Source: https://arxiv.org/abs/2406.19395
- Identifier: arXiv:2406.19395
- Released: 2024-06-27
- First featured: Quant Letter No. 55 (2024-07-03): https://www.ml-quant.com/issues/2024-07-03/
- Citations (Semantic Scholar): 9
- Published in: not yet
- Topic: Other

## Related

- [Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions](https://www.ml-quant.com/papers/arxiv/2411.14405/): The research investigates the OpenAI o1 model's ability, enhanced by Chain-of-Thought fine-tuning and innovative reasoning strategies, to adapt to wider domains lacking clear standards.
- [NV-Retriever: Improving text embedding models with effective hard-negative mining](https://www.ml-quant.com/papers/arxiv/2407.15831/): Hard-Negative Mining: The article suggests positive-aware mining methods for fine-tuning text embedding models, with the NV-Retriever-v1 model outperforming previous methods in the MTEB Retrieval benchmark.
- [LoRA-X: Bridging Foundation Models with Training-Free Cross-Model Adaptation](https://www.ml-quant.com/papers/arxiv/2501.16559/): Model Adaptation: LoRA-X enables the transfer of fine-tuning parameters across different models, enhancing the efficiency of text-to-image generation without needing original training data.
- [Exploring Foundation Models for Synthetic Medical Imaging: A Study on Chest X-Rays and Fine-Tuning Techniques](https://www.ml-quant.com/papers/arxiv/2409.04424/): The study investigates the use of foundation models in creating realistic medical images, specifically chest x-rays, and shows performance improvement with fine-tuning.
- [Instruction Finetuning Llama3-8B Model Using LoRA for Financial Named Entity Recognition](https://www.ml-quant.com/papers/arxiv/2601.10043/): The paper shows that using instruction fine-tuning and Low-Rank Adaptation with Meta's Llama 3 enhances financial named-entity recognition, leading to top performance in converting unformatted reports into organized knowledge.
- [Depth Anything V2](https://www.ml-quant.com/papers/arxiv/2406.09414/): Depth Anything V2 is a new model for monocular depth estimation, using synthetic and large-scale pseudo-labeled real images for faster, more accurate results and setting a new evaluation benchmark.
