preg_match_all(PHP)中的UTF-8字符 [英] UTF-8 characters in preg_match_all (PHP)

查看：168 发布时间：2020/5/27 3:02:09 php regex pcre

本文介绍了preg_match_all(PHP)中的UTF-8字符的处理方法，对大家解决问题具有一定的参考价值，需要的朋友们下面随着小编来一起学习吧！

问题描述

我有preg_match_all('/[aäeëioöuáéíóú]/u', $in, $out, PREG_OFFSET_CAPTURE);

如果$in = 'hëllo' $out是:

array(1) {
[0]=>
  array(2) {
  [0]=>
    array(2) {
      [0]=>
      string(2) "ë"
  [1]=>
  int(1)
}
[1]=>
array(2) {
  [0]=>
  string(1) "o"
  [1]=>
  int(5)
  }
}
}

o的位置应为4.我已在线阅读有关此问题的信息(ë被计为2).有解决方案吗?我见过mb_substr和类似的内容，但是preg_match_all是否有类似的内容?

The position of o should be 4. I've read about this problem online (the ë gets counted as 2). Is there a solution for this? I've seen mb_substr and similar, but is there something like this for preg_match_all?

相关种类:它们在Python中是否等同于preg_match_all? (返回匹配项及其在字符串中的位置)

Kind of related: Is their an equivalent of preg_match_all in Python? (Returning an array of matches with their position in the string)

preg_match_all(PHP)中的UTF-8字符 [英] UTF-8 characters in preg_match_all (PHP)

问题描述

推荐答案

相关文章

PHP最新文章

热门教程

热门工具

登录关闭

preg_match_all(PHP)中的UTF-8字符 [英] UTF-8 characters in preg_match_all (PHP)

问题描述

推荐答案

相关文章

PHP最新文章

热门教程

热门工具

登录 关闭

登录关闭