Solr简介

最新推荐文章于 2020-09-25 09:17:02 发布

坦GA

最新推荐文章于 2020-09-25 09:17:02 发布

阅读量617

点赞数

分类专栏： Lucene/Solr 文章标签： solr

Lucene/Solr 专栏收录该内容

9 篇文章 1 订阅

订阅专栏

原文地址：http://zh.hortonworks.com/apache/solr/

Rapid indexing & search on Hadoop

Apache Solr is the open source platform for searches of data stored in HDFS in Hadoop. Solr powers the search and navigation features of many of the world’s largest Internet sites, enabling powerful full-text search and near real-time indexing. Whether users search for tabular（表格式的）, text, geo-location or sensor（传感式） data in Hadoop, they find it quickly with Apache Solr.

WHAT SOLR DOES

Hadoop operators put documents in Apache Solr by “indexing” via XML, JSON, CSV or binary over HTTP.

Then users can query those petabytes of data via HTTP GET. They can receive XML, JSON, CSV or binary results. Apache Solr is optimized for high volume web traffic.

Top features include:

Advanced full-text search
Near real-time indexing
Standards-based open interfaces like XML, JSON and HTTP
Comprehensive HTML administration interfaces
Server statistics exposed over JMX for monitoring
Linearly scalable, auto index replication, auto failover and recovery
Flexible and adaptable, with XML configuration

Solr is highly reliable, scalable and fault tolerant. Both data analysts and developers in the open source community trust Solr’s distributed indexing, replication and load-balanced querying capabilities.