• <xmp id="1ykh9"><source id="1ykh9"><mark id="1ykh9"></mark></source></xmp>
      <b id="1ykh9"><small id="1ykh9"></small></b>
    1. <b id="1ykh9"></b>

      1. <button id="1ykh9"></button>
        <video id="1ykh9"></video>
      2. west china medical publishers
        Keyword
        • Title
        • Author
        • Keyword
        • Abstract
        Advance search
        Advance search

        Search

        find Keyword "Retrieval-augmented generation" 1 results
        • Performance evaluation of lightweight Chinese large language models integrated with retrieval-augmented generation technology in answering specialized lung cancer questions

          Objective To evaluate the performance of lightweight Chinese large language models (LLMs) in answering specialized lung cancer questions, and to explore the impact of retrieval-augmented generation (RAG) on model performance. Methods Eleven lightweight Chinese LLMs with parameter sizes ranging from 7B to 32B were included. A lung cancer-specific evaluation dataset consisting of 200 questions [100 A1-type (basic knowledge) and 100 A2-type (clinical case) questions], constructed based on clinical guidelines and thoracic surgery textbooks, was used for assessment. Model performance was evaluated under two conditions (with and without RAG). Accuracy was used to assess model performance, and response latency was recorded to reflect inference efficiency. An accuracy–latency scatter plot was constructed for descriptive analysis of overall model performance. Results All models successfully completed the evaluation. With the introduction of RAG, the overall average accuracy improved from 61.68% to 76.36%. Smaller models demonstrated the most significant improvement (e.g., the accuracy of DeepSeek-7B increased from 32.50% to 60.00%, P<0.001). The average response latency increased from 12.58 s to 13.80 s. The Qwen3 series showed the best overall performance, and Qwen3-32B achieved the highest accuracy under both conditions (76.50% and 84.00%, respectively). After RAG integration, performance differences among model families were markedly reduced. Based on the accuracy-latency trade-off, Qwen3-32B achieved the best balance between accuracy and response latency under the baseline condition, whereas Qwen3-14B demonstrated superior overall performance in terms of accuracy, latency, and computational cost after RAG integration. Conclusion The integration of RAG technology improves the ability of lightweight Chinese LLMs to answer specialized lung cancer questions. Under the dual practical constraints of limited computational resources and medical data security requirements, the "lightweight model+RAG" technical framework may represent a promising deployment solution.

          Release date: Export PDF Favorites Scan
        1 pages Previous 1 Next

        Format

        Content

      3. <xmp id="1ykh9"><source id="1ykh9"><mark id="1ykh9"></mark></source></xmp>
          <b id="1ykh9"><small id="1ykh9"></small></b>
        1. <b id="1ykh9"></b>

          1. <button id="1ykh9"></button>
            <video id="1ykh9"></video>
          2. 射丝袜