Version at: 16/02/2016, 07:28 vs. version at: 20/05/2016, 22:35
11#How to Search for Text
22
33Return to [tatoeba.org](https://tatoeba.org/eng/sentences/advanced_search).
44
55
66## Important Note
77
88The search engine on tatoeba.org doesn't work like other standard search engines.
99
1010You can't use ? or ! in your searches in the way you would normally expect to use them, so you need to search for sentences without using these punctuation marks.
1111
1212## Tatoeba.org uses [Sphinx Search](http://sphinxsearch.com/docs/current.html#boolean-syntax)
1313
1414These instructions tell you how to use the search bar at the top of every Tatoeba page. Our search works much like a search engine such as Google, but has some important differences.
1515
1616* To find English sentences with "live", "lives", "living" or "lived", search for the word "live". (This will also find sentences with "Live", "Living", etc., since capitalization is ignored.)
1717
1818 * [live](http://tatoeba.org/eng/sentences/search?query=live+&from=eng&to=und)
1919
2020* To match a word exactly (ignoring capitalization), put an equals sign (=) before it.
2121
2222 * [=live](http://tatoeba.org/eng/sentences/search?query=%3Dlive+&from=eng&to=und)
2323
2424* Leave punctuation out of your search string. Most punctuation will be ignored, but a final exclamation mark (!) or question mark (?) will actually interfere with the search. These symbols have other purposes, as described later on this page.
2525 * The following yields no results:
2626
2727 * [how strange!](http://tatoeba.org/eng/sentences/search?query=how+strange!&from=eng&to=und)
2828
2929 * but this search will find *How strange!* among other results:
3030
3131 * [how strange](http://tatoeba.org/eng/sentences/search?query=how+strange&from=eng&to=und)
3232
3333* Put a $ after a word to find sentences ending with that word. The example finds English sentences ending with "Tom".
3434
3535 * [Tom$](http://tatoeba.org/eng/sentences/search?query=Tom%24&from=eng&to=und)
36
37* Note that because $ is a special character, if you would like to search for sentences containing the symbol $, you will need to escape the symbol with a backslash.
38
39 * [\$](https://tatoeba.org/eng/sentences/search?query=\%24&from=und&to=und)
3640
3741* Put a ^ before a word to find sentences beginning with that word. The example finds English sentences beginning with "Tom".
3842
3943 * [^Tom](http://tatoeba.org/eng/sentences/search?query=%5ETom&from=eng&to=und)
4044
4145* This example finds English sentences beginning with "Tom" and ending with "Mary".
4246
4347 * [^Tom Mary$](http://tatoeba.org/eng/sentences/search?query=%5ETom+Mary%24&from=eng&to=und)
4448
4549* This example finds English sentences beginning with either "Tom" or "He".
4650
4751 * [(^Tom|^He)](http://tatoeba.org/eng/sentences/search?query=%28%5ETom%7C%5EHe%29&from=eng&to=und)
4852
4953
5054* To search for a phrase, put quotes (") around it. Put an equals sign in front of each word that you want to be matched exactly.
5155 * If you want to see phrases like "live in Boston", "living in Boston", or "lives in Boston", use the following search:
5256
5357 * ["live in boston"](http://tatoeba.org/eng/sentences/search?query=%22live+in+boston%22&from=eng&to=und)
5458
5559 * The following search will only find sentences with the exact phrase "live in Boston".
5660
5761 * ["=live =in =boston"](http://tatoeba.org/eng/sentences/search?query=%22%3Dlive+%3Din+%3Dboston%22&from=eng&to=und)
5862
5963 * This search will only find sentences consisting of the exact words "I live in Boston"..
6064
6165 * ["^I =live =in =Boston$"](http://tatoeba.org/eng/sentences/search?query=%22%5EI+%3Dlive+%3Din+%3DBoston%24%22&from=eng&to=und)
6266
6367* This example finds English sentences that have "Tom", but don't begin with "Tom."
6468
6569 * [-^Tom Tom](https://tatoeba.org/eng/sentences/search?query=-%5ETom+Tom&from=eng&to=und)
6670
6771* This example finds English sentences that have "Tom", but don't begin or end with "Tom."
6872
6973 * [-^Tom Tom -Tom$](https://tatoeba.org/eng/sentences/search?query=-%5ETom+Tom+-Tom%24&from=eng&to=und)
7074
7175* The question mark (?) as part of a word is a one-letter wildcard.
7276
7377 * The following will find sentences with either "whenever" and "wherever."
7478
7579 * [whe?ever](https://tatoeba.org/eng/sentences/search?query=whe%3Fever&from=und&to=und)
7680
7781 * The following will find sentences with with 6-letter words that have 2 letters, and then "eve" and then one more letter, such as "clever" "eleven", "peeves", "uneven", ...
7882
7983 * [??eve?](https://tatoeba.org/eng/sentences/search?query=%3F%3Feve%3F&from=eng&to=und)
8084
8185
8286* This example finds English sentences that have "Tom", then 2 words, then "Mary", then 1 word, and then "John."
8387
8488 * ["Tom * * Mary * John"](https://tatoeba.org/eng/sentences/search?query=%22Tom+*+*+Mary+*+John%22&from=eng&to=und)
8589
8690* This example finds English sentences that start with "Tom", then 3 words, then ends with "Mary".
8791
8892 * ["^Tom * * * Mary$"](https://tatoeba.org/eng/sentences/search?query=%22%5ETom+*+*+*+Mary%24%22&from=und&to=und)
8993
9094
9195
9296* This example finds English sentences that have words beginning with "red", including the word "red". (3 letters or more are required.)
9397
9498 * [red*](https://tatoeba.org/eng/sentences/search?query=red*&from=eng&to=und)
9599
96100* This example finds English sentences that have words ending with "red", including the word "red".
97101
98102 * [*red](https://tatoeba.org/eng/sentences/search?query=*red&from=eng&to=und)
99103
100104* This example finds English sentences that have words containing the word "red", including the word "red".
101105
102106 * [\*red\*](https://tatoeba.org/eng/sentences/search?query=*red*&from=eng&to=und)
103107
104108* This example finds English sentences that have the word "French", but don't have the word "Tom"
105109
106110 * [French -Tom](https://tatoeba.org/eng/sentences/search?query=French+-Tom&from=eng&to=und)
107111
108112* This example will find sentences with "cheek" (in any form: cheeks, etc.) that don't include any of the words preceded by a minus sign (-).
109113
110114 * [cheek -tear -slap -burn -red -hollow](http://tatoeba.org/eng/sentences/search?query=cheek+-tear+-slap+-burn+-red+-hollow&from=eng&to=und)
111115
112116* This example finds sentences in which the word "cat" comes before the word "dog."
113117
114118 * [cat << dog](https://tatoeba.org/eng/sentences/search?query=cat+%3C%3C+dog&from=eng&to=und)
115119
116120## Languages without word boundaries
117121
118122For languages that don't use space characters to separate words, like Japanese, Chinese etc. the search engine interprets each character as a single word. For instance, searching for 逆に will return the same results as 逆 に, which actually matches sentences that only *include* these characters, but not necessarily in that particular order, or not contiguously. So you want to surround keywords with quotes: ["逆に"](http://tatoeba.org/jpn/sentences/search?query=%22%E9%80%86%E3%81%AB%22&from=jpn).
119123
120124## More details
121125
122126The search ignores capitalization and punctuation (unless the punctuation happens to match one of the special characters described elsewhere on the page). An apostrophe within a word is not treated as punctuation, so you can find such words as "don't" by including them in an ordinary search string.
123127
124128In some languages, including English, the search engine **stems** the search words by default. This means that it removes certain trailing sequences from both search words and indexed words. Thus a search for *live* will also find *lived* and *living*.
125129
126130The languages in which the search engine stems words are: German, English, Finnish, French, Italian, Dutch, Portuguese, Russian, Spanish, Swedish and Turkish.
127131
128132If you want to find an exact match for a word, you must precede it with an equals sign, as in *=live*. This may come as a surprise to users who are accustomed to Google Search, where wrapping a word or phrase in double quotes forces an exact match. In Sphinx, double quotes have a different function, which only affects multiword (phrase) searches: wrapping a phrase in double quotes requires matching sentences to contain words in the specified continuous sequence. Simply placing a phrase in quotes does not suppress stemming of its individual words. To do that, you will need to place an equals sign before each word in the phrase for which you want to suppress stemming.
129133
130134As an example, take the search *like thing*. This will find *like things*, *likely things*, and even *things like*. Adding quotes, as in *"like thing"*, will prevent a match against *things like* (where the words appear in the wrong order), but it will continue to match *like things*, *likely things*, and so on. By contrast, *"=like =thing"* will only match *like thing* (which does not occur in the Tatoeba corpus). Removing the double quotes, *=like =thing*, will match *What made you do a silly thing like that?* Removing one of the equals signs, as in *like =thing*, will find *Such a strange thing is not likely to happen.*
131135
132136Note that a star (*) can be placed at the beginning and/or end of a string representing a word, but it if is placed in the middle, the search will always fail. Also, a string beginning and/or ending with a star must be at least three characters long.
133137
134138## Other search operators
135139
136140* A vertical bar (representing "or") finds examples where either of the words appears:
137141 * *hate | detest* will match sentences with either *hate* or *detest* (or both).
138142
139143* If you want to combine an or-expression with other terms, you need to put the or-expression in parentheses:
140144 * *(red|blue) house* will match sentences in which the word "house" appears together with either "red" or "blue" (or both)
141145
142146* A dash (or exclamation point) before a word prevents matches with sentences where the word appears: *like -thing* (or *like !thing*) will match *I like ice cream* but not *I like that red thing*.
143147
144148* Putting a caret (^) before a word will match only sentences that begin with that word: *^great* will match *Great people are not always wise.* but not *You are the great love of my life.*
145149
146150* Putting a dollar sign ($) after a word will match only sentences that end with that word: *life$* will match *This is the best day of my life.* but not *Life means nothing without friends.*
147151
148152* If you want to search for sentences that contain nothing other than the specified words, use double quotes, a caret, and a dollar sign in combination: *"^i love you$"* will find *I love you.* and *I love you!* but not *I love you more than you love me.* (However, it will find *I loved you.* To prevent this match, use *"^i =love you$"*.)
149153
150154* The strict order operator (<<) between two words will find sentences where the first word occurs before the second but not where the second word comes before the first. Thus _dog << cat_ will find examples where _dog_ precedes _cat_, but not vice versa.
151155
152156See the [Sphinx documentation](http://sphinxsearch.com/docs/current.html#boolean-syntax) for other functionality. Note that the documentation mentions keywords pertaining to specific fields in a document, but these are not relevant to Tatoeba.
153157
154158
diff view generated by jsdifflib

Version at: 16/02/2016, 07:28

#How to Search for Text

Return to [tatoeba.org](https://tatoeba.org/eng/sentences/advanced_search).


## Important Note

The search engine on tatoeba.org doesn't work like other standard search engines.

You can't use ? or ! in your searches in the way you would normally expect to use them, so you need to search for sentences without using these punctuation marks.

## Tatoeba.org uses  [Sphinx Search](http://sphinxsearch.com/docs/current.html#boolean-syntax) 

These instructions tell you how to use the search bar at the top of every Tatoeba page. Our search works much like a search engine such as Google, but has some important differences. 

* To find English sentences with "live", "lives", "living" or "lived", search for the word "live". (This will also find sentences with "Live", "Living", etc., since capitalization is ignored.)

  * [live](http://tatoeba.org/eng/sentences/search?query=live+&from=eng&to=und)

* To match a word exactly (ignoring capitalization), put an equals sign (=) before it. 

  * [=live](http://tatoeba.org/eng/sentences/search?query=%3Dlive+&from=eng&to=und)

* Leave punctuation out of your search string. Most punctuation will be ignored, but a final exclamation mark (!) or question mark (?) will actually interfere with the search. These symbols have other purposes, as described later on this page.
  * The following yields no results:

      * [how strange!](http://tatoeba.org/eng/sentences/search?query=how+strange!&from=eng&to=und)

  * but this search will find *How strange!* among other results:

      * [how strange](http://tatoeba.org/eng/sentences/search?query=how+strange&from=eng&to=und)

* Put a $ after a word to find sentences ending with that word. The example finds English sentences ending with "Tom".

  * [Tom$](http://tatoeba.org/eng/sentences/search?query=Tom%24&from=eng&to=und)

* Put a ^ before  a word to find sentences beginning with that word. The example finds English sentences beginning with "Tom".

  * [^Tom](http://tatoeba.org/eng/sentences/search?query=%5ETom&from=eng&to=und)

* This example finds English sentences beginning with "Tom" and ending with "Mary".

  * [^Tom Mary$](http://tatoeba.org/eng/sentences/search?query=%5ETom+Mary%24&from=eng&to=und)

* This example finds English sentences beginning with either "Tom" or "He".

  * [(^Tom|^He)](http://tatoeba.org/eng/sentences/search?query=%28%5ETom%7C%5EHe%29&from=eng&to=und)


* To search for a phrase, put quotes (") around it. Put an equals sign in front of each word that you want to be matched exactly.
  * If you want to see phrases like "live in Boston", "living in Boston", or "lives in Boston", use the following search:

      * ["live in boston"](http://tatoeba.org/eng/sentences/search?query=%22live+in+boston%22&from=eng&to=und)

  * The following search will only find sentences with the exact phrase "live in Boston".

      * ["=live =in =boston"](http://tatoeba.org/eng/sentences/search?query=%22%3Dlive+%3Din+%3Dboston%22&from=eng&to=und)

  * This search will only find sentences consisting of the exact words "I live in Boston"..

      * ["^I =live =in =Boston$"](http://tatoeba.org/eng/sentences/search?query=%22%5EI+%3Dlive+%3Din+%3DBoston%24%22&from=eng&to=und)

* This example finds English sentences that have "Tom", but don't begin with "Tom."

  * [-^Tom Tom](https://tatoeba.org/eng/sentences/search?query=-%5ETom+Tom&from=eng&to=und)

* This example finds English sentences that have "Tom", but don't begin or end with "Tom."

  * [-^Tom Tom -Tom$](https://tatoeba.org/eng/sentences/search?query=-%5ETom+Tom+-Tom%24&from=eng&to=und)

* The question mark (?) as part of a word is a one-letter wildcard.

    * The following will find sentences with either "whenever" and "wherever."

        * [whe?ever](https://tatoeba.org/eng/sentences/search?query=whe%3Fever&from=und&to=und)

    * The following will find sentences with with 6-letter words that have 2 letters, and then "eve" and then one more letter,  such as "clever" "eleven", "peeves", "uneven", ...

        * [??eve?](https://tatoeba.org/eng/sentences/search?query=%3F%3Feve%3F&from=eng&to=und)


* This example finds English sentences that have "Tom", then 2 words, then "Mary", then 1 word, and then "John."

  * ["Tom * * Mary * John"](https://tatoeba.org/eng/sentences/search?query=%22Tom+*+*+Mary+*+John%22&from=eng&to=und)

* This example finds English sentences that start with "Tom", then 3 words, then ends with "Mary".

  * ["^Tom * * * Mary$"](https://tatoeba.org/eng/sentences/search?query=%22%5ETom+*+*+*+Mary%24%22&from=und&to=und)



* This example finds English sentences that have words beginning with "red", including the word "red".  (3 letters or more are required.)

  * [red*](https://tatoeba.org/eng/sentences/search?query=red*&from=eng&to=und)

* This example finds English sentences that have words ending with "red", including the word "red".

  * [*red](https://tatoeba.org/eng/sentences/search?query=*red&from=eng&to=und)

* This example finds English sentences that have words containing the word "red", including the word "red".

  * [\*red\*](https://tatoeba.org/eng/sentences/search?query=*red*&from=eng&to=und)

* This example finds English sentences that have the word "French", but don't have the word "Tom"

  * [French -Tom](https://tatoeba.org/eng/sentences/search?query=French+-Tom&from=eng&to=und)

* This example will find sentences with "cheek" (in any form: cheeks, etc.) that don't include any of the words preceded by a minus sign (-).

  * [cheek -tear -slap -burn -red -hollow](http://tatoeba.org/eng/sentences/search?query=cheek+-tear+-slap+-burn+-red+-hollow&from=eng&to=und)

* This example finds sentences in which the word "cat" comes before the word "dog."

  * [cat << dog](https://tatoeba.org/eng/sentences/search?query=cat+%3C%3C+dog&from=eng&to=und)

## Languages without word boundaries

For languages that don't use space characters to separate words, like Japanese, Chinese etc. the search engine interprets each character as a single word. For instance, searching for 逆に will return the same results as 逆 に, which actually matches sentences that only *include* these characters, but not necessarily in that particular order, or not contiguously. So you want to surround keywords with quotes: ["逆に"](http://tatoeba.org/jpn/sentences/search?query=%22%E9%80%86%E3%81%AB%22&from=jpn).

## More details

The search ignores capitalization and punctuation (unless the punctuation happens to match one of the special characters described elsewhere on the page). An apostrophe within a word is not treated as punctuation, so you can find such words as "don't" by including them in an ordinary search string. 

In some languages, including English, the search engine **stems** the search words by default. This means that it removes certain trailing sequences from both search words and indexed words. Thus a search for *live* will also find *lived* and *living*.

The languages in which the search engine stems words are: German, English, Finnish, French, Italian, Dutch, Portuguese, Russian, Spanish, Swedish and Turkish.

If you want to find an exact match for a word, you must precede it with an equals sign, as in *=live*. This may come as a surprise to users who are accustomed to Google Search, where wrapping a word or phrase in double quotes forces an exact match. In Sphinx, double quotes have a different function, which only affects multiword (phrase) searches: wrapping a phrase in double quotes requires matching sentences to contain words in the specified continuous sequence. Simply placing a phrase in quotes does not suppress stemming of its individual words. To do that, you will need to place an equals sign before each word in the phrase for which you want to suppress stemming.

As an example, take the search *like thing*. This will find *like things*, *likely things*, and even *things like*. Adding quotes, as in *"like thing"*, will prevent a match against *things like* (where the words appear in the wrong order), but it will continue to match *like things*, *likely things*, and so on. By contrast, *"=like =thing"* will only match *like thing* (which does not occur in the Tatoeba corpus). Removing the double quotes, *=like =thing*, will match *What made you do a silly thing like that?* Removing one of the equals signs, as in *like =thing*, will find *Such a strange thing is not likely to happen.* 

Note that a star (*) can be placed at the beginning and/or end of a string representing a word, but it if is placed in the middle, the search will always fail. Also, a string beginning and/or ending with a star must be at least three characters long.

## Other search operators

* A vertical bar (representing "or") finds examples where either of the words appears:
  *    *hate | detest* will match sentences with either *hate* or *detest* (or both). 

* If you want to combine an or-expression with other terms, you need to put the or-expression in parentheses: 
  *    *(red|blue) house* will match sentences in which the word "house" appears together with either "red" or "blue" (or both) 

* A dash (or exclamation point) before a word prevents matches with sentences where the word appears: *like -thing* (or *like !thing*) will match *I like ice cream* but not *I like that red thing*.

* Putting a caret (^) before a word will match only sentences that begin with that word: *^great* will match *Great people are not always wise.* but not *You are the great love of my life.* 

* Putting a dollar sign ($) after a word will match only sentences that end with that word: *life$* will match *This is the best day of my life.* but not *Life means nothing without friends.*

* If you want to search for sentences that contain nothing other than the specified words, use double quotes, a caret, and a dollar sign in combination: *"^i love you$"* will find *I love you.* and *I love you!* but not *I love you more than you love me.* (However, it will find *I loved you.* To prevent this match, use *"^i =love you$"*.)

* The strict order operator (<<) between two words will find sentences where the first word occurs before the second but not where the second word comes before the first. Thus _dog << cat_ will find examples where _dog_ precedes _cat_, but not vice versa.

See the [Sphinx documentation](http://sphinxsearch.com/docs/current.html#boolean-syntax) for other functionality. Note that the documentation mentions keywords pertaining to specific fields in a document, but these are not relevant to Tatoeba.

version at: 20/05/2016, 22:35

#How to Search for Text

Return to [tatoeba.org](https://tatoeba.org/eng/sentences/advanced_search).


## Important Note

The search engine on tatoeba.org doesn't work like other standard search engines.

You can't use ? or ! in your searches in the way you would normally expect to use them, so you need to search for sentences without using these punctuation marks.

## Tatoeba.org uses  [Sphinx Search](http://sphinxsearch.com/docs/current.html#boolean-syntax) 

These instructions tell you how to use the search bar at the top of every Tatoeba page. Our search works much like a search engine such as Google, but has some important differences. 

* To find English sentences with "live", "lives", "living" or "lived", search for the word "live". (This will also find sentences with "Live", "Living", etc., since capitalization is ignored.)

  * [live](http://tatoeba.org/eng/sentences/search?query=live+&from=eng&to=und)

* To match a word exactly (ignoring capitalization), put an equals sign (=) before it. 

  * [=live](http://tatoeba.org/eng/sentences/search?query=%3Dlive+&from=eng&to=und)

* Leave punctuation out of your search string. Most punctuation will be ignored, but a final exclamation mark (!) or question mark (?) will actually interfere with the search. These symbols have other purposes, as described later on this page.
  * The following yields no results:

      * [how strange!](http://tatoeba.org/eng/sentences/search?query=how+strange!&from=eng&to=und)

  * but this search will find *How strange!* among other results:

      * [how strange](http://tatoeba.org/eng/sentences/search?query=how+strange&from=eng&to=und)

* Put a $ after a word to find sentences ending with that word. The example finds English sentences ending with "Tom".

  * [Tom$](http://tatoeba.org/eng/sentences/search?query=Tom%24&from=eng&to=und)

* Note that because $ is a special character, if you would like to search for sentences containing the symbol $, you will need to escape the symbol with a backslash.

  * [\$](https://tatoeba.org/eng/sentences/search?query=\%24&from=und&to=und)

* Put a ^ before  a word to find sentences beginning with that word. The example finds English sentences beginning with "Tom".

  * [^Tom](http://tatoeba.org/eng/sentences/search?query=%5ETom&from=eng&to=und)

* This example finds English sentences beginning with "Tom" and ending with "Mary".

  * [^Tom Mary$](http://tatoeba.org/eng/sentences/search?query=%5ETom+Mary%24&from=eng&to=und)

* This example finds English sentences beginning with either "Tom" or "He".

  * [(^Tom|^He)](http://tatoeba.org/eng/sentences/search?query=%28%5ETom%7C%5EHe%29&from=eng&to=und)


* To search for a phrase, put quotes (") around it. Put an equals sign in front of each word that you want to be matched exactly.
  * If you want to see phrases like "live in Boston", "living in Boston", or "lives in Boston", use the following search:

      * ["live in boston"](http://tatoeba.org/eng/sentences/search?query=%22live+in+boston%22&from=eng&to=und)

  * The following search will only find sentences with the exact phrase "live in Boston".

      * ["=live =in =boston"](http://tatoeba.org/eng/sentences/search?query=%22%3Dlive+%3Din+%3Dboston%22&from=eng&to=und)

  * This search will only find sentences consisting of the exact words "I live in Boston"..

      * ["^I =live =in =Boston$"](http://tatoeba.org/eng/sentences/search?query=%22%5EI+%3Dlive+%3Din+%3DBoston%24%22&from=eng&to=und)

* This example finds English sentences that have "Tom", but don't begin with "Tom."

  * [-^Tom Tom](https://tatoeba.org/eng/sentences/search?query=-%5ETom+Tom&from=eng&to=und)

* This example finds English sentences that have "Tom", but don't begin or end with "Tom."

  * [-^Tom Tom -Tom$](https://tatoeba.org/eng/sentences/search?query=-%5ETom+Tom+-Tom%24&from=eng&to=und)

* The question mark (?) as part of a word is a one-letter wildcard.

    * The following will find sentences with either "whenever" and "wherever."

        * [whe?ever](https://tatoeba.org/eng/sentences/search?query=whe%3Fever&from=und&to=und)

    * The following will find sentences with with 6-letter words that have 2 letters, and then "eve" and then one more letter,  such as "clever" "eleven", "peeves", "uneven", ...

        * [??eve?](https://tatoeba.org/eng/sentences/search?query=%3F%3Feve%3F&from=eng&to=und)


* This example finds English sentences that have "Tom", then 2 words, then "Mary", then 1 word, and then "John."

  * ["Tom * * Mary * John"](https://tatoeba.org/eng/sentences/search?query=%22Tom+*+*+Mary+*+John%22&from=eng&to=und)

* This example finds English sentences that start with "Tom", then 3 words, then ends with "Mary".

  * ["^Tom * * * Mary$"](https://tatoeba.org/eng/sentences/search?query=%22%5ETom+*+*+*+Mary%24%22&from=und&to=und)



* This example finds English sentences that have words beginning with "red", including the word "red".  (3 letters or more are required.)

  * [red*](https://tatoeba.org/eng/sentences/search?query=red*&from=eng&to=und)

* This example finds English sentences that have words ending with "red", including the word "red".

  * [*red](https://tatoeba.org/eng/sentences/search?query=*red&from=eng&to=und)

* This example finds English sentences that have words containing the word "red", including the word "red".

  * [\*red\*](https://tatoeba.org/eng/sentences/search?query=*red*&from=eng&to=und)

* This example finds English sentences that have the word "French", but don't have the word "Tom"

  * [French -Tom](https://tatoeba.org/eng/sentences/search?query=French+-Tom&from=eng&to=und)

* This example will find sentences with "cheek" (in any form: cheeks, etc.) that don't include any of the words preceded by a minus sign (-).

  * [cheek -tear -slap -burn -red -hollow](http://tatoeba.org/eng/sentences/search?query=cheek+-tear+-slap+-burn+-red+-hollow&from=eng&to=und)

* This example finds sentences in which the word "cat" comes before the word "dog."

  * [cat << dog](https://tatoeba.org/eng/sentences/search?query=cat+%3C%3C+dog&from=eng&to=und)

## Languages without word boundaries

For languages that don't use space characters to separate words, like Japanese, Chinese etc. the search engine interprets each character as a single word. For instance, searching for 逆に will return the same results as 逆 に, which actually matches sentences that only *include* these characters, but not necessarily in that particular order, or not contiguously. So you want to surround keywords with quotes: ["逆に"](http://tatoeba.org/jpn/sentences/search?query=%22%E9%80%86%E3%81%AB%22&from=jpn).

## More details

The search ignores capitalization and punctuation (unless the punctuation happens to match one of the special characters described elsewhere on the page). An apostrophe within a word is not treated as punctuation, so you can find such words as "don't" by including them in an ordinary search string. 

In some languages, including English, the search engine **stems** the search words by default. This means that it removes certain trailing sequences from both search words and indexed words. Thus a search for *live* will also find *lived* and *living*.

The languages in which the search engine stems words are: German, English, Finnish, French, Italian, Dutch, Portuguese, Russian, Spanish, Swedish and Turkish.

If you want to find an exact match for a word, you must precede it with an equals sign, as in *=live*. This may come as a surprise to users who are accustomed to Google Search, where wrapping a word or phrase in double quotes forces an exact match. In Sphinx, double quotes have a different function, which only affects multiword (phrase) searches: wrapping a phrase in double quotes requires matching sentences to contain words in the specified continuous sequence. Simply placing a phrase in quotes does not suppress stemming of its individual words. To do that, you will need to place an equals sign before each word in the phrase for which you want to suppress stemming.

As an example, take the search *like thing*. This will find *like things*, *likely things*, and even *things like*. Adding quotes, as in *"like thing"*, will prevent a match against *things like* (where the words appear in the wrong order), but it will continue to match *like things*, *likely things*, and so on. By contrast, *"=like =thing"* will only match *like thing* (which does not occur in the Tatoeba corpus). Removing the double quotes, *=like =thing*, will match *What made you do a silly thing like that?* Removing one of the equals signs, as in *like =thing*, will find *Such a strange thing is not likely to happen.* 

Note that a star (*) can be placed at the beginning and/or end of a string representing a word, but it if is placed in the middle, the search will always fail. Also, a string beginning and/or ending with a star must be at least three characters long.

## Other search operators

* A vertical bar (representing "or") finds examples where either of the words appears:
  *    *hate | detest* will match sentences with either *hate* or *detest* (or both). 

* If you want to combine an or-expression with other terms, you need to put the or-expression in parentheses: 
  *    *(red|blue) house* will match sentences in which the word "house" appears together with either "red" or "blue" (or both) 

* A dash (or exclamation point) before a word prevents matches with sentences where the word appears: *like -thing* (or *like !thing*) will match *I like ice cream* but not *I like that red thing*.

* Putting a caret (^) before a word will match only sentences that begin with that word: *^great* will match *Great people are not always wise.* but not *You are the great love of my life.* 

* Putting a dollar sign ($) after a word will match only sentences that end with that word: *life$* will match *This is the best day of my life.* but not *Life means nothing without friends.*

* If you want to search for sentences that contain nothing other than the specified words, use double quotes, a caret, and a dollar sign in combination: *"^i love you$"* will find *I love you.* and *I love you!* but not *I love you more than you love me.* (However, it will find *I loved you.* To prevent this match, use *"^i =love you$"*.)

* The strict order operator (<<) between two words will find sentences where the first word occurs before the second but not where the second word comes before the first. Thus _dog << cat_ will find examples where _dog_ precedes _cat_, but not vice versa.

See the [Sphinx documentation](http://sphinxsearch.com/docs/current.html#boolean-syntax) for other functionality. Note that the documentation mentions keywords pertaining to specific fields in a document, but these are not relevant to Tatoeba.

Note

The lines in green are the lines that have been added in the new version. The lines in red are those that have been removed.