Skip to content

REGEX module

This module offers matching operations against regular expressions using the powerful PCRE library.

A text file containing regular expressions categorized in groups is compiled when the module is loaded, storing the compiled PCRE objects in an array. A function to match a string or pseudo-variable against any of these groups is provided. The text file can be modified and reloaded at any time via a MI command. The module also offers a function to perform a PCRE matching operation against a regular expression provided as function parameter.

For a detailed list of PCRE features read the man page of the library.

The following modules must be loaded before this module:

  • No dependencies on other OpenSIPS modules.

The following libraries or applications must be installed before running OpenSIPS with this module loaded:

  • libpcre-dev - the development libraries of PCRE.

Text file containing the regular expression groups. It must be set in order to enable the group matching function.

Default value is “NULL”.

Set file parameter
...
modparam("regex", "file", "/etc/opensips/regex_groups")
...

Max number of regular expression groups in the text file.

Default value is “20”.

Set max_groups parameter
...
modparam("regex", "max_groups", 40)
...

Max content size of a group in the text file.

Default value is “8192”.

Set group_max_size parameter
...
modparam("regex", "group_max_size", 16384)
...

If this options is set, matching is done caseless. It is equivalent to Perl’s /i option, and it can be changed within a pattern by a (?i) or (?-i) option setting.

Default value is “0”.

Set pcre_caseless parameter
...
modparam("regex", "pcre_caseless", 1)
...

By default, PCRE treats the subject string as consisting of a single line of characters (even if it actually contains newlines). The “start of line” metacharacter (^) matches only at the start of the string, while the “end of line” metacharacter ($) matches only at the end of the string, or before a terminating newline.

When this option is set, the “start of line” and “end of line” constructs match immediately following or immediately before internal newlines in the subject string, respectively, as well as at the very start and end. This is equivalent to Perl’s /m option, and it can be changed within a pattern by a (?m) or (?-m) option setting. If there are no newlines in a subject string, or no occurrences of ^ or $ in a pattern, setting this option has no effect.

Default value is “0”.

Set pcre_multiline parameter
...
modparam("regex", "pcre_multiline", 1)
...

If this option is set, a dot metacharater in the pattern matches all characters, including those that indicate newline. Without it, a dot does not match when the current position is at a newline. This option is equivalent to Perl’s /s option, and it can be changed within a pattern by a (?s) or (?-s) option setting.

Default value is “0”.

Set pcre_dotall parameter
...
modparam("regex", "pcre_dotall", 1)
...

If this option is set, whitespace data characters in the pattern are totally ignored except when escaped or inside a character class. Whitespace does not include the VT character (code 11). In addition, characters between an unescaped # outside a character class and the next newline, inclusive, are also ignored. This is equivalent to Perl’s /x option, and it can be changed within a pattern by a (?x) or (?-x) option setting.

Default value is “0”.

Set pcre_extended parameter
...
modparam("regex", "pcre_extended", 1)
...

Matches the given string parameter against the regular expression pcre_regex, which is compiled into a PCRE object. Returns TRUE if it matches, FALSE otherwise.

Meaning of the parameters is as follows:

  • string - String or pseudo-variable to compare.
  • pcre_regex - Regular expression to be compiled in a PCRE object. It can be a string or pseudo-variable.

This function can be used from REQUEST_ROUTE, FAILURE_ROUTE, ONREPLY_ROUTE, BRANCH_ROUTE and LOCAL_ROUTE.

pcre_match usage (forcing case insensitive)
...
if (pcre_match("$ua", "(?i)^twinkle")) {
xlog("L_INFO", "User-Agent matches\n");
}
...
pcre_match usage (using 'end of line' symbol)
...
if (pcre_match("$rU", "^user[1234]$$")) { # Will be converted to "^user[1234]$"
xlog("L_INFO", "RURI username matches\n");
}
...

It uses the groups readed from the text file (see file format id) to match the given string parameter against the compiled regular expression in group number group. Returns TRUE if it matches, FALSE otherwise.

Meaning of the parameters is as follows:

  • string - String or pseudo-variable to compare.
  • group - Number of group to use in the operation. If not specified then 0 (the first group) is used.

This function can be used from REQUEST_ROUTE, FAILURE_ROUTE, ONREPLY_ROUTE, BRANCH_ROUTE and LOCAL_ROUTE.

pcre_match_group usage
...
if (pcre_match_group("$rU", "2")) {
xlog("L_INFO", "RURI username matches group 2\n");
}
...

Causes regex module to re-read the content of the text file and re-compile the regular expressions. The number of groups in the file can be modified safely.

Name: regex_reload

Parameters: none

MI FIFO Command Format:

Terminal window
:regex_reload:_reply_fifo_file_
_empty_line_

The file contains regular expressions categorized in groups. Each group starts with “[number]” line. Lines starting by space, tab, CR, LF or # (comments) are ignored. Each regular expression must take up just one line, this means that a regular expression can’t be splitted in various lines.

An example of the file format would be the following:

regex file
### List of User-Agents publishing presence status
[0]
# Softphones
^Twinkle/1
^X-Lite
^eyeBeam
^Bria
^SIP Communicator
^Linphone
# Deskphones
^Snom
# Others
^SIPp
^PJSUA
### Blacklisted source IP's
[1]
^190\.232\.250\.226$
^122\.5\.27\.125$
^86\.92\.112\.
### Free PSTN destinations in Spain
[2]
^1\d{3}$
^((\+|00)34)?900\d{6}$

The module compiles the text above to the following regular expressions:

group 0: ((^Twinkle/1)|(^X-Lite)|(^eyeBeam)|(^Bria)|(^SIP Communicator)|
(^Linphone)|(^Snom)|(^SIPp)|(^PJSUA))
group 1: ((^190\.232\.250\.226$)|(^122\.5\.27\.125$)|(^86\.92\.112\.))
group 2: ((^1\d{3}$)|(^((\+|00)34)?900\d{6}$))

The first group can be used to avoid auto-generated PUBLISH (pua_usrloc module) for UA’s already supporting presence:

Using with pua_usrloc
route[REGISTER] {
if (! pcre_match_group("$ua", 0)) {
xlog("L_INFO", "Auto-generated PUBLISH for $fu ($ua)\n");
pua_set_publish();
}
save("location");
exit;
}
Incorrect groups file
[1]
^aaa
^bbb
[2]
^ccc
^ddd

will generate the following regular expressions:

group 0: ((^aaa)|(^bbb))
group 1: ((^ccc)|(^ddd))
[0]
([0-9]{9})
( #abcde)
( qwerty)

Top contributors by DevScore(1), authored commits(2) and lines added/removed(3)

#NameDevScoreCommitsLines++Lines—
1.Iñaki Baz Castillo15312422
2.Stefan Darius (@dariusstefan)12550281
3.Razvan Crainea (@razvancrainea)971411
4.Liviu Chircu (@liviuchircu)856884
5.Bogdan-Andrei Iancu (@bogdan-iancu)751511
6.Sergio Gutierrez31202
7.Ovidiu Sas (@ovidiusas)311311
8.Anca Vamanu3163
9.Marius Zbihlei2120

(1) DevScore = author_commits + author_lines_added / (project_lines_added / project_commits) + author_lines_deleted / (project_lines_deleted / project_commits)

(2) including any documentation-related commits, excluding merge commits

(3) ignoring whitespace edits, renamed files and auto-generated files

#NameCommit Activity
1.Stefan Darius (@dariusstefan)Jun 2026 - Jul 2026
2.Razvan Crainea (@razvancrainea)Sep 2011 - Jun 2026
3.Liviu Chircu (@liviuchircu)Mar 2014 - Jun 2018
4.Bogdan-Andrei Iancu (@bogdan-iancu)Feb 2009 - Oct 2014
5.Ovidiu Sas (@ovidiusas)Jan 2013 - Jan 2013
6.Marius ZbihleiSep 2010 - Sep 2010
7.Iñaki Baz CastilloFeb 2009 - Jul 2010
8.Anca VamanuSep 2009 - Sep 2009
9.Sergio GutierrezFeb 2009 - Feb 2009

(1) including any documentation-related commits, excluding merge commits

Last edited by: Razvan Crainea (@razvancrainea), Liviu Chircu (@liviuchircu), Bogdan-Andrei Iancu (@bogdan-iancu), Iñaki Baz Castillo.

All documentation files (i.e. .md extension) are licensed under the Creative Common License 4.0