# Pattern tokenization to split multiple URL's (edited)

**URL:** <https://discuss.elastic.co/t/pattern-tokenization-to-split-multiple-urls-edited/32525>\
**Category:** Elasticsearch\
**Created:** [October 19, 2015, 7:19pm UTC](https://discuss.elastic.co/t/pattern-tokenization-to-split-multiple-urls-edited/32525 "2015-10-19T19:19:00Z")\
**Posts on this page:** 2\
**Page:** 1

<div class="post-metadata">

**Author:** ![Phrozyn](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/phrozyn/32/5207_2.png) [@Phrozyn](https://discuss.elastic.co/u/Phrozyn)\
**Post date:** [October 19, 2015, 7:19pm UTC](https://discuss.elastic.co/t/pattern-tokenization-to-split-multiple-urls-edited/32525/1 "2015-10-19T19:19:00Z")

</div>

I've got a field that is parsing on every non-alphanumeric character, and I'd like to change how it's being parsed to only on comma's.

I've been struggling with how to use the pattern tokenizer.

The entries in this field are generally FQDNs separated by commas like: [www.domain.com](http://www.domain.com),[blah-blah.domain.com](http://blah-blah.domain.com),[some.domain.com](http://some.domain.com)

My analyzer mapping looks like this:  
{"settings":{"analysis":{"analyzer":{"comma":{"type":"pattern","pattern":"\,+"}}}}}

my field mapping looks like this:  
{"type":"string", "analysis": {"analyzer":{"comma": {"tokenizer": "pattern"}}}},"os\_family":{"type":"string"},

so now I get [www.domain.com](http://www.domain.com) but [blah-blah.domain.com](http://blah-blah.domain.com) is getting separated at blah and blah

I have no hyphen in my pattern, any ideas?

Thank you.

---

<div class="post-metadata">

**Author:** ![system](https://us1.discourse-cdn.com/elastic/original/3X/1/a/1ac57faf039f6b580b3f104ef42a2a89e41014de.png) [@system](https://discuss.elastic.co/u/system)\
**Post date:** [July 5, 2017, 11:43pm UTC](https://discuss.elastic.co/t/pattern-tokenization-to-split-multiple-urls-edited/32525/2 "2017-07-05T23:43:49Z")

</div>


