zoukankan      html  css  js  c++  java
  • [转]solr系统query检索词特殊字符的处理

    原文地址:http://blog.csdn.net/wgw335363240/article/details/39889979

    solr是基于 lucence开发的应用,如果query中带有非法字符串,结果很可能是检索出所有内容或者直接报错,所以你对用户的输入必须要先做处理。输入星号,能够检索出所有内容;输入加号,则会报错。

    官方的处理办法(java,因为solr是java开发的):

    https://svn.apache.org/repos/asf/lucene/dev/trunk/solr/solrj/src/java/org/apache/solr/client/solrj/util/ClientUtils.java
    
    public static String escapeQueryChars(String s) {
        StringBuilder sb = new StringBuilder();
        for (int i = 0; i < s.length(); i++) {
          char c = s.charAt(i);
          // These characters are part of the query syntax and must be escaped
          if (c == '\' || c == '+' || c == '-' || c == '!'  || c == '(' || c == ')' || c == ':'
            || c == '^' || c == '[' || c == ']' || c == '"' || c == '{' || c == '}' || c == '~'
            || c == '*' || c == '?' || c == '|' || c == '&'  || c == ';' || c == '/'
            || Character.isWhitespace(c)) {
            sb.append('\');
          }
          sb.append(c);
        }
        return sb.toString();
      }

    翻译的php版本(利用preg_replace函数进行正则替换):

    static public function escape($value)
    {
        //list taken from http://lucene.apache.org/java/docs/queryparsersyntax.html#Escaping%20Special%20Characters
        $pattern = '/(+|-|&|||!|(|)|{|}|[|]|^|"|~|*|?|:|;|~|/)/';
        $replace = '\$1';
    
       return preg_replace($pattern, $replace, $value);
    }

    翻译后的python版本:

    import re
    def escape_solr(word):
        return re.sub('(\|+|-|&||||!|(|)|{|}|[|]|^|"|~|*|?|:|;|/|~)','\1', word )
  • 相关阅读:
    SpringBoot 之 静态资源路径、显示首页、错误页
    微擎框架的缓存机制实现源码解读
    SpringBoot 之 多环境切换
    SpringBoot 之 JSR303 数据校验
    CSS——NO.6(盒模型)
    CSS——NO.5(格式化排版)
    CSS——NO.4(继承、层叠、特殊性、重要性)
    CSS——NO.3(CSS选择器)
    CSS——NO.2(CSS样式的基本知识)
    CSS——NO.1(初识CSS)
  • 原文地址:https://www.cnblogs.com/fashflying/p/6638391.html
Copyright © 2011-2022 走看看