TDM
TDM(Text and Data Mining) refers to automated processes using computer software to gather, extract, and analyze large volumes of academic text and data.
If you need to automatically collect and analyze large numbers of research papers for academic purposes, please use the official TDM methods provided by publishers (such as APIs or dedicated request procedures) rather than repeatedly downloading, crawling, or scraping content directly from publisher websites.
※ Permission to conduct TDM does not necessarily permit the use of content for AI model training, generative AI services, or RAG applications. Such uses may be subject to separate publisher licenses and terms of use.
If you need to automatically collect and analyze large numbers of research papers for academic purposes, please use the official TDM methods provided by publishers (such as APIs or dedicated request procedures) rather than repeatedly downloading, crawling, or scraping content directly from publisher websites.
※ Permission to conduct TDM does not necessarily permit the use of content for AI model training, generative AI services, or RAG applications. Such uses may be subject to separate publisher licenses and terms of use.
TDM Guidelines by Publisher
| Publisher | Access Method | Guide |
|---|---|---|
| Issue API Key from Developer Portal | ||
| Agree to the TDM Terms and obtain an API Token | ||
| Apply via Data Solutions → Case-by-case publisher agreement → Issue API Key | ||
| Apply via Library → Case-by-case publisher agreement → Data access within permitted range | ||
| Apply via Library → Data access within permitted range |
ACS Guidelines
- Application: Submit applicant information (Affiliation, Status, Name, Email), TDM purpose, and the DOI list of target materials to the library(libres@korea.ac.kr)
- Agreement: The library and publisher will enter into a case-by-case TDM agreement based on the submitted details.
- Data Access: Once the agreement is executed, download and analyze data within the permitted scope.
- Important Notes – Delivery Method: APIs are not provided; users must download PDF full-texts manually.
– Restrictions: Use outside TDM scope, AI/LLM training, public dissemination of outputs, or joint research with external organizations may be restricted.
– Approval & Fees: Coverage and potential fees depend on publisher review and approval (Additional fees may apply for requests exceeding 20,000 items).
RSC Guidelines
- Application: Submit applicant information (Affiliation, Status, Name, Email), TDM purpose, start/end dates, target DOI list, and crawling IP address to the library (libres@korea.ac.kr)
- Data Access: Download and analyze data within the permitted scope.
- Important Notes – Delivery Method: Provided in HTML format. XML data requires separate purchase.
– Restrictions: Uses outside TDM purposes may be restricted.
– Approval & Fees: Additional fees apply for requests exceeding 2,000 items
General Precautions for TDM Use
- Access is limited strictly to content for which the institution holds valid subscription rights.
- Automated bulk downloading via website crawling or scraping is strictly prohibited.
- Sharing authentication credentials (API Keys, Tokens, etc.) with third parties is prohibited.
- Posting or redistributing downloaded full-texts to public websites, GitHub, or external repositories is prohibited.
- Users must comply with each publisher’s terms regarding data retention, deletion, and request limits.
Notice on AI Use of E-Resources
Subscribed content from major academic publishers may be subject to specific restrictions or conditions regarding AI model training, the use of generative AI tools, and Text and Data Mining (TDM).
Please be sure to check the publisher’s license terms and AI-related usage conditions before use. Any violation of the terms of use may result in an institution-wide suspension of access to electronic resources.
Please exercise particular caution when using electronic resources for research and study.
- Check whether the target content is subscribed to by the library.
- Verify whether the AI service reuses or trains on user input data, or shares it with third parties.
- Review publisher terms based on your intended AI usage (e.g., summarization/analysis, model training, RAG, chatbot integration).
Publisher AI Guidelines Direct Links
FAQ
- Q: Can I bulk download papers if it is strictly for research purposes?
A: Research purposes alone do not permit unrestricted bulk downloading. If you need to collect a large number of articles automatically, please use the TDM methods officially provided by the publisher, such as an API or a separate approval process. - Q: Can I upload paper PDFs to generative AI tools like ChatGPT to summarize them?
A: Please check the publisher’s license terms and the AI service’s terms of use before doing so. The use of Library-subscribed electronic resources with generative AI tools may be restricted under the terms of the subscription agreement. Please review the publisher’s AI-related usage terms before using such content. - Q: Can I collect thousands of papers for RAG or AI model training?
A: TDM authorization alone does not permit the use of content for AI model training or RAG. The use of Library-subscribed electronic resources with generative AI tools may be restricted under the terms of the subscription agreement. Please review the publisher’s AI-related usage terms before using such content.