<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Analysis on INFINI Pizza</title><link>/docs/references/search/analysis/</link><description>Recent content in Analysis on INFINI Pizza</description><generator>Hugo</generator><language>en</language><atom:link href="/docs/references/search/analysis/index.xml" rel="self" type="application/rss+xml"/><item><title>Analyze text</title><link>/docs/references/search/analysis/analysis/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>/docs/references/search/analysis/analysis/</guid><description>&lt;h1 id="analyze-text">
 Analyze text
 &lt;a class="anchor" href="#analyze-text">#&lt;/a>
&lt;/h1>
&lt;p>The analysis workbench: run the server&amp;rsquo;s tokenizers, normalizers, token
filters and analyzers over sample text and see the exact token stream —
the same components the indexing path uses, so what you debug here is
exactly what collections run.&lt;/p>
&lt;h2 id="list-components">
 List components
 &lt;a class="anchor" href="#list-components">#&lt;/a>
&lt;/h2>
&lt;div class="highlight">&lt;pre tabindex="0" class="chroma">&lt;code class="language-sh" data-lang="sh">&lt;span class="line">&lt;span class="cl">GET /_analysis/components
&lt;/span>&lt;/span>&lt;/code>&lt;/pre>&lt;/div>&lt;p>Returns every registered tokenizer, normalizer, token filter and named
analyzer (the set a schema may reference — see

 &lt;a href="/docs/references/collection/create/">custom analyzers&lt;/a> for defining per-collection
ones). The server ships &lt;strong>83 analyzers, 50 tokenizers, 15 normalizers
and 327 token filters&lt;/strong> out of the box — the full catalog is on the

 &lt;a href="/docs/references/search/analysis/analysis-builtins/">built-in components&lt;/a> page:&lt;/p></description></item><item><title>Built-in analysis components</title><link>/docs/references/search/analysis/analysis-builtins/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>/docs/references/search/analysis/analysis-builtins/</guid><description>&lt;h1 id="built-in-analysis-components">
 Built-in analysis components
 &lt;a class="anchor" href="#built-in-analysis-components">#&lt;/a>
&lt;/h1>
&lt;p>The server ships &lt;strong>83&lt;/strong> analyzers, &lt;strong>50&lt;/strong> tokenizers,
&lt;strong>15&lt;/strong> normalizers and &lt;strong>327&lt;/strong> token filters, registered
at startup. The authoritative, always-current list on YOUR node is:&lt;/p>
&lt;div class="highlight">&lt;pre tabindex="0" class="chroma">&lt;code class="language-sh" data-lang="sh">&lt;span class="line">&lt;span class="cl">GET /_analysis/components
&lt;/span>&lt;/span>&lt;/code>&lt;/pre>&lt;/div>&lt;p>Components are referenced by name — in a field&amp;rsquo;s schema
(&lt;code>&amp;quot;analyzer&amp;quot;: &amp;quot;english&amp;quot;&lt;/code>), in a 
 &lt;a href="/docs/references/collection/create/">custom analyzer
definition&lt;/a>, or in an ad-hoc

 &lt;a href="/docs/references/search/analysis/analysis/">&lt;code>_analysis/analyze&lt;/code>&lt;/a> pipeline.&lt;/p>
&lt;h2 id="analyzers">
 Analyzers
 &lt;a class="anchor" href="#analyzers">#&lt;/a>
&lt;/h2>
&lt;p>General purpose:&lt;/p>
&lt;p>&lt;code>fingerprint&lt;/code>, &lt;code>keyword&lt;/code>, &lt;code>pattern&lt;/code>, &lt;code>simple&lt;/code>, &lt;code>standard&lt;/code>, &lt;code>stop&lt;/code>, &lt;code>whitespace&lt;/code>&lt;/p>
&lt;p>CJK and transliteration:&lt;/p>
&lt;p>&lt;code>cjk&lt;/code>, &lt;code>ik_max_word&lt;/code>, &lt;code>ik_smart&lt;/code>, &lt;code>jieba&lt;/code>, &lt;code>kuromoji&lt;/code>, &lt;code>nori&lt;/code>, &lt;code>pinyin&lt;/code>, &lt;code>smartcn&lt;/code>
&lt;code>stconvert_s2t&lt;/code>, &lt;code>stconvert_t2s&lt;/code>&lt;/p></description></item></channel></rss>