Bibliomining
Bibliomining is the use of a combination of data mining, data warehousing, and bibliometrics for the purpose of analyzing library services.[1][2] The term was created in 2003 by Scott Nicholson, Assistant Professor, Syracuse University School of Information Studies, in order to distinguish data mining in a library setting from other types of data mining.[3]
How bibliomining works
First a data warehouse must be created. This is done by compiling information on the resources, such as titles and authors, subject headings, and descriptions of the collections. Then the demographic surrogate information is organized. Finally the library information (such as the librarian, whether or not the information came from the reference desk or circulation desk, and the location of the library) is obtained.
Once this is organized, the data can be processed and analyzed. This can be done via a few methods, such as online analytical processing (OLAP), using a data mining program, or through data visualization.
Uses of bibliomining
Bibliomining is used to discover patterns in what people are reading and researching and allows librarians to target their community better. Bibliomining can also help library directors focus their budgets on resources that will be utilized. Another use is to determine when people use the library more often, so staffing needs can be adequately met. Combining bibliomining with other research techniques such as focus groups, surveys and cost-benefit analysis, will help librarians to get a better picture of their patrons and their needs.
Issues
There is some concern that data mining violates patron privacy. But by extracting the data, all personally identifiable information is deleted, and the data warehouse is clean. The original patron data can then be totally deleted and there will be no way to link the new data to a particular patron. This can be done in a few ways. One, used with information regarding database access, is to track the IP address, but then replace it with a similar code, that will allow identification without violating privacy. Another is to keep track of an item returned to the library and create a "demographic surrogate" of the patron. The demographic surrogate would not give any identifiable information such as names, library card numbers or addresses.
The other concern in bibliomining is that it only provides data in a very detached manner. Information is given as to how a patron uses library resources, but there is no way to track if the resources met the user's needs completely. Someone could take out a book on a topic, but not find the information they were seeking. Bibliomining only helps identify which books are used, not how useful they actually were. Bibliomining cannot provide information on how well a collection serves a patron. In order to counteract this, bibliomining must be used in accordance with other research techniques.
See also
References
- ^ Jiann, Cherng Shieh (2010). "The integration system for librarians' bibliomining". The Electronic Library. 28 (5). Emerald Group Publishing Limited: 709–721. doi:10.1108/02640471011081988. hdl:2241/102576. S2CID 27191290.
- ^ Nicholson, Scott (May 2006). "The basis for bibliomining: Frameworks for bringing together usage-based data mining and bibliometrics through data warehousing in digital library services". Information Processing and Management. 42 (3): 785–804. doi:10.1016/j.ipm.2005.05.008. hdl:10150/106175.
- ^ Nicholson, Scott. "Bibliomining for Library Decision-Making". ResearchGate. Retrieved 1 January 2006.
Further reading
- Nicholson, S. (2003). The Bibliomining Process: Data Warehousing and Data Mining for Library Decision-Making Information Technology and Libraries 22 (4), 146-151.
- Gunther, K. (2000). Applying data mining principles to a library data collection — Data mining can help you make decisions and serve patrons better. Computers in Libraries 20(4), 60-63.
Content Disclaimer
Informasi ini disarikan dari Wikipedia dan disajikan kembali untuk tujuan edukasi. Konten tersedia di bawah lisensi CC BY-SA 3.0. Kami tidak bertanggung jawab atas ketidakakuratan data yang bersumber dari kontribusi publik tersebut.
- The information displayed on this website is sourced in part or in whole from Wikipedia and has been adapted for the purpose of restating it. We strive to provide accurate and relevant information, however:
- There is no guarantee of absolute accuracy. Wikipedia is an open, collaborative project that can be edited by anyone, so information is subject to change.
- It is not intended to constitute professional advice. The content displayed is for informational and educational purposes only. For important decisions (e.g., medical, legal, or financial), please consult a professional.
- Content copyright. Wikipedia is licensed under the Creative Commons Attribution-ShareAlike License (CC BY-SA). This means that content may be reused with appropriate attribution and shared under a similar license.
- Responsible use. Any risk arising from the use of information from this website is entirely the responsibility of the user.