Text Mining5.9一款用于文本挖掘的软件

软件来源微信公众号【学术点滴】

【1】Text Mining5.9中文版软件核心功能:

(1)多个文本自定义分词

频次统计

词云图绘制

主题聚类

(2)单个大文本自定义分词

频次统计

词云图绘制

主题聚类

(3)突破VOS软件只能做数据库数据的局限

(4)使用VOSviewer软件做任意网络文本的图谱

(5)基于TF-IDF算法的多文本关键词提取

(6)基于textrank算法的多文本关键词提取

(7)用户评论情感分析

(8)基于词袋模型的LDA主题挖掘

(9)基于TF-IDF模型的LDA主题挖掘

(10)神经网络语言模型

(11)关键词去重

(12)同义词一键合并

(13)无意义词一键删除

【1】Text Mining5.9英文版软件核心功能:

(1)多个文本自定义分词

频次统计

词云图绘制

主题聚类

(2)单个大文本自定义分词

频次统计

词云图绘制

主题聚类

(3)该软件同时可突破VOS软件只能做数据库数据的局限

(4)使用VOSviewer软件做任意网络文本的图谱

(5)基于TF-IDF算法的多文本关键词提取

(6)基于词袋模型的LDA主题挖掘

(7)基于TF-IDF模型的LDA主题挖掘

(8)神经网络语言模型

(9)关键词去重

(10)同义词一键合并

(11)无意义词一键删除

(12)英语词组提取
在这里插入图片描述
在这里插入图片描述

  • 0
    点赞
  • 5
    收藏
    觉得还不错? 一键收藏
  • 0
    评论
Key Features Develop all the relevant skills for building text-mining apps with R with this easy-to-follow guide Gain in-depth understanding of the text mining process with lucid implementation in the R language Example-rich guide that lets you gain high-quality information from text data Book Description Text Mining (or text data mining or text analytics) is the process of extracting useful and high-quality information from text by devising patterns and trends. R provides an extensive ecosystem to mine text through its many frameworks and packages. Starting with basic information about the statistics concepts used in text mining, this book will teach you how to access, cleanse, and process text using the R language and will equip you with the tools and the associated knowledge about different tagging, chunking, and entailment approaches and their usage in natural language processing. Moving on, this book will teach you different dimensionality reduction techniques and their implementation in R. Next, we will cover pattern recognition in text data utilizing classification mechanisms, perform entity recognition, and develop an ontology learning framework. By the end of the book, you will develop a practical application from the concepts learned, and will understand how text mining can be leveraged to analyze the massively available data on social media. What you will learn Get acquainted with some of the highly efficient R packages such as OpenNLP and RWeka to perform various steps in the text mining process Access and manipulate data from different sources such as JSON and HTTP Process text using regular expressions Get to know the different approaches of tagging texts, such as POS tagging, to get started with text analysis Explore different dimensionality reduction techniques, such as Principal Component Analysis (PCA), and understand its implementation in R Discover the underlying themes or topics that are present in an unstructured collection of documents, using common topic models such as Latent Dirichlet Allocation (LDA) Build a baseline sentence completing application Perform entity extraction and named entity recognition using R About the Author Ashish Kumar is an IIM alumnus and an engineer at heart. He has extensive experience in data science, machine learning, and natural language processing having worked at organizations, such as McAfee-Intel, an ambitious data science startup Volt consulting), and presently associated to the software and research lab of a leading MNC. Apart from work, Ashish also participates in data science competitions at Kaggle in his spare time. Avinash Paul is a programming language enthusiast, loves exploring open sources technologies and programmer by choice. He has over nine years of programming experience. He has worked in Sabre Holdings , McAfee , Mindtree and has experience in data-driven product development, He was intrigued by data science and data mining while developing niche product in education space for a ambitious data science start-up. He believes data science can solve lot of societal challenges. In his spare time he loves to read technical books and teach underprivileged children back home. Table of Contents Chapter 1. Statistical Linguistics with R Chapter 2. Processing Text Chapter 3. Categorizing and Tagging Text Chapter 4. Dimensionality Reduction Chapter 5. Text Summarization and Clustering Chapter 6. Text Classification Chapter 7. Entity Recognition

“相关推荐”对你有帮助么?

  • 非常没帮助
  • 没帮助
  • 一般
  • 有帮助
  • 非常有帮助
提交
评论
添加红包

请填写红包祝福语或标题

红包个数最小为10个

红包金额最低5元

当前余额3.43前往充值 >
需支付:10.00
成就一亿技术人!
领取后你会自动成为博主和红包主的粉丝 规则
hope_wisdom
发出的红包
实付
使用余额支付
点击重新获取
扫码支付
钱包余额 0

抵扣说明:

1.余额是钱包充值的虚拟货币,按照1:1的比例进行支付金额的抵扣。
2.余额无法直接购买下载,可以购买VIP、付费专栏及课程。

余额充值