*lucene索引_的删除和更新

【删除】

【恢复删除】

【强制删除】

【优化和合并】

【更新索引】

附：

代码：

IndexUtil.java：

 package cn.hk.index;

 import java.io.File;

 import java.io.IOException;

 import org.apache.lucene.analysis.standard.StandardAnalyzer;

 import org.apache.lucene.document.Document;

 import org.apache.lucene.document.Field;

 import org.apache.lucene.index.CorruptIndexException;

 import org.apache.lucene.index.IndexReader;

 import org.apache.lucene.index.IndexWriter;

 import org.apache.lucene.index.IndexWriterConfig;

 import org.apache.lucene.index.StaleReaderException;

 import org.apache.lucene.index.Term;

 import org.apache.lucene.store.Directory;

 import org.apache.lucene.store.FSDirectory;

 import org.apache.lucene.store.LockObtainFailedException;

 import org.apache.lucene.util.Version;

 public class IndexUtil {

     private String[] ids = {"1","2","3","4","5","6"};

     private String[] emails = {"aa@hk.arg","bb@hk.org","cc@hk.arg",

                                "dd@hk.org","ee@hk.org","ff@hk.org"};

     private String[] content = {

             "welcome to visited the space","hello boy","my name is aa","i like football",

             "I like football and I like Basketball too","I like movie and swim"

     };

     private int[] attachs = {2,3,1,4,5,5};

     private String[] names = {"zhangsan","lisi","john","mike","jetty","jake"};

     private Directory directory = null;

     public IndexUtil(){

         try {

             directory = FSDirectory.open(new File("d://lucene/index02"));

         } catch (IOException e) {

             e.printStackTrace();

         }

     }

     public void update(){

         IndexWriter writer =null;

         try {

             writer = new IndexWriter(directory,

                     new IndexWriterConfig(Version.LUCENE_35,new StandardAnalyzer(Version.LUCENE_35)));

             /*

              * lucene并没有提供更新的方法，这里的更新其实是提供如下两个操作：

              * 先删除之后再添加

              */

             Document doc = new Document();

             doc.add(new Field("id","11",Field.Store.YES,Field.Index.NOT_ANALYZED_NO_NORMS));

             doc.add(new Field("email",emails[0],Field.Store.YES,Field.Index.NOT_ANALYZED));

             doc.add(new Field("content",content[0],Field.Store.NO,Field.Index.ANALYZED));

             doc.add(new Field("name",names[0],Field.Store.YES,Field.Index.NOT_ANALYZED_NO_NORMS));

             writer.updateDocument(new Term("id","1"),doc);

         } catch (CorruptIndexException e) {

             e.printStackTrace();

         } catch (LockObtainFailedException e) {

             e.printStackTrace();

         } catch (IOException e) {

             e.printStackTrace();

         }finally{

             if(writer != null)

                 try {

                     writer.close();

                 } catch (CorruptIndexException e) {

                     e.printStackTrace();

                 } catch (IOException e) {

                     e.printStackTrace();

                 }

         }

     }

     public void merge(){

         IndexWriter writer = null;

         try {

             writer = new IndexWriter(directory,

                     new IndexWriterConfig(Version.LUCENE_35, new StandardAnalyzer(Version.LUCENE_35)));

             //会将索引合并为2段，这两段中的被删除的数据会被清空

             //特别注意：此处在lucene3.5后不建议使用，因为会消耗大量的开销，

             //lucene会根据情况自动处理的

             writer.forceMerge(2);

         } catch (CorruptIndexException e) {

             e.printStackTrace();

         } catch (LockObtainFailedException e) {

             e.printStackTrace();

         } catch (IOException e) {

             e.printStackTrace();

         }finally{

             if(writer != null)

                 try {

                     writer.close();

                 } catch (CorruptIndexException e) {

                     e.printStackTrace();

                 } catch (IOException e) {

                     e.printStackTrace();

                 }

         }

     }

     public void forceDelete(){

         IndexWriter writer = null;

         try {

             writer = new IndexWriter(directory,

                     new IndexWriterConfig(Version.LUCENE_35,new StandardAnalyzer(Version.LUCENE_35)));

             writer.forceMergeDeletes();

         } catch (CorruptIndexException e) {

             e.printStackTrace();

         } catch (LockObtainFailedException e) {

             e.printStackTrace();

         } catch (IOException e) {

             e.printStackTrace();

         }finally{

             if(writer != null)

                 try {

                     writer.close();

                 } catch (CorruptIndexException e) {

                     e.printStackTrace();

                 } catch (IOException e) {

                     e.printStackTrace();

                 }

         }

     }

     public void undelete(){

         //使用IndexReader进行恢复

         try {

             IndexReader reader = IndexReader.open(directory,false);

             //回复时，必须把IndexReader的只读（readyonly）设置为FALSE

             reader.undeleteAll();

             reader.close();

         } catch (StaleReaderException e) {

             e.printStackTrace();

         } catch (CorruptIndexException e) {

             e.printStackTrace();

         } catch (LockObtainFailedException e) {

             e.printStackTrace();

         } catch (IOException e) {

             // TODO Auto-generated catch block

             e.printStackTrace();

         }

     }

     public void delete(){

         IndexWriter writer = null;

         try {

             writer = new IndexWriter(directory,

                     new IndexWriterConfig(Version.LUCENE_35, new StandardAnalyzer(Version.LUCENE_35)));

             //删除ID为1的文档

             //参数可以是一个选项，可以是一个Query，也可以是一个Term，Term是一个精确查找的值

             //此时删除的文档并不会被完全删除，而是存储在回收站中的，可以恢复

             writer.deleteDocuments(new Term("id","1"));

         } catch (CorruptIndexException e) {

             e.printStackTrace();

         } catch (LockObtainFailedException e) {

             e.printStackTrace();

         } catch (IOException e) {

             e.printStackTrace();

         }finally{

             if(writer != null)

                 try {

                     writer.close();

                 } catch (CorruptIndexException e) {

                     e.printStackTrace();

                 } catch (IOException e) {

                     e.printStackTrace();

                 }

         }

     }

     public void query(){

         try {

             IndexReader reader = IndexReader.open(directory);

             //通过reader可以获取文档的数量

             System.out.println("numDocs:" + reader.numDocs());

             System.out.println("maxDocs" + reader.maxDoc());

             System.out.println("deleteDocs:" + reader.numDeletedDocs());

             reader.close();

         } catch (CorruptIndexException e) {

             e.printStackTrace();

         } catch (IOException e) {

             e.printStackTrace();

         }

     }

     public void index(){

         IndexWriter writer = null;

         try {

             writer = new IndexWriter(directory,new IndexWriterConfig(Version.LUCENE_35, new StandardAnalyzer(Version.LUCENE_35)));

             Document doc = null;

             for(int i=0;i<ids.length;i++){

                 doc = new Document();

                 doc.add(new Field("id",ids[i],Field.Store.YES,Field.Index.NOT_ANALYZED_NO_NORMS));

                 doc.add(new Field("email",emails[i],Field.Store.YES,Field.Index.NOT_ANALYZED));

                 doc.add(new Field("content",content[i],Field.Store.NO,Field.Index.ANALYZED));

                 doc.add(new Field("name",names[i],Field.Store.YES,Field.Index.NOT_ANALYZED_NO_NORMS));

                 writer.addDocument(doc);

             }

         } catch (CorruptIndexException e) {

             e.printStackTrace();

         } catch (LockObtainFailedException e) {

             e.printStackTrace();

         } catch (IOException e) {

             e.printStackTrace();

         }finally{

                 if(writer != null)

                     try {

                         writer.close();

                     } catch (CorruptIndexException e) {

                         e.printStackTrace();

                     } catch (IOException e) {

                         e.printStackTrace();

                     }

         }

     }

 }

TestIndex.java：

 package cn.hk.test;

 import org.junit.Test;

 import cn.hk.index.IndexUtil;

 public class TestIndex {

     @Test

     public void testIndex(){

         IndexUtil iu = new IndexUtil();

         iu.index();

     }

     @Test

     public void testQuery(){

         IndexUtil iu = new IndexUtil();

         iu.query();

     }

     @Test

     public void testDelete(){

         IndexUtil iu = new IndexUtil();

         iu.delete();

     }

     @Test

     public void testUnDelete(){

         IndexUtil iu = new IndexUtil();

         iu.undelete();

     }

     @Test

     public void testForceDelete(){

         IndexUtil iu = new IndexUtil();

         iu.forceDelete();

     }

     public void testMerge(){

         IndexUtil  iu = new IndexUtil();

         iu.merge();

     }

     @Test

     public void testUpdate(){

         IndexUtil iu = new IndexUtil();

         iu.update();

     }

 }

*lucene索引_的删除和更新的更多相关文章

*lucene索引_创建_域选项
[索引建立步骤] [创建Directory] [创建writer] [创建文档并添加索引] 文档和域的概念很重要文档相当于表中的每一条记录,域相当于表中的每一个字段. [查询索引的基本信息] 使用I ...
Lucene系列五：Lucene索引详解（IndexWriter详解、Document详解、索引更新）
一.IndexWriter详解问题1:索引创建过程完成什么事? 分词.存储到反向索引中 1. 回顾Lucene架构图: 介绍我们编写的应用程序要完成数据的收集,再将数据以document的形式用lu ...
Lucene——索引的创建、删除、修改
package cn.tz.lucene; import java.io.File; import java.util.ArrayList; import java.util.List; import ...
Lucene索引维护(添加、修改、删除)
1. Field域属性分类添加文档的时候,我们文档当中包含多个域,那么域的类型是我们自定义的,上个案例使用的TextField域,那么这个域他会自动分词,然后存储我们要根据数 ...
Lucene——索引过程分析Index
Lucene索引过程分为3个主要操作步骤:将原始文档转换成文本.分析文本.将分析好的文本保存至索引中一.提取文本和创建文档从 pdf.word等非纯文本格式文件中,提取文本格式信息.建立起对应的, ...
【手把手教你全文检索】Lucene索引的【增、删、改、查】
前言搞检索的,应该多少都会了解Lucene一些,它开源而且简单上手,官方API足够编写些小DEMO.并且根据倒排索引,实现快速检索.本文就简单的实现增量添加索引,删除索引,通过关键字查询,以及更新索 ...
理解Lucene索引与搜索过程中的核心类
理解索引过程中的核心类执行简单索引的时候需要用的类有: IndexWriter.Directory.Analyzer.Document.Field 1.IndexWriter IndexWr ...
lucene索引
一.lucene索引 1.文档层次结构索引(Index):一个索引放在一个文件夹中: 段(Segment):一个索引中可以有很多段,段与段之间是独立的,添加新的文档可能产生新段,不同的段可以合并成一 ...
Lucene 索引功能
Lucene 数据建模基本概念文档(doc): 文档是 Lucene 索引和搜索的原子单元,文档是一个包含多个域的容器. 域(field): 域包含“真正的”被搜索的内容,每一个域都有一个标识名称 ...

随机推荐

bzoj 4513 [Sdoi2016]储能表
题面 https://www.lydsy.com/JudgeOnline/problem.php?id=4513 题解要求的式子用数位dp的方法去做我们把式子拆开变成 $\sum_{i=0}^ ...
[BZOJ5120]无限之环
Description 曾经有一款流行的游戏,叫做InfinityLoop,先来简单的介绍一下这个游戏: 游戏在一个n×m的网格状棋盘上进行,其中有些小方格中会有水管,水管可能在方格某些方向的边界的中 ...
【模板】c++动态数组vector
相信大家都知道$C$++里有一个流弊的$STL$模板库.. 今天我们就要谈一谈这里面的一个容器:动态数组$vector$. $vector$实际上类似于$a[]$这个东西,也就是说它重载了$[]$运算 ...
线段树(单点更新) HDU 1754 I Hate It
题目传送门 /* 线段树基本功能:区间最大值,修改某个值 */ #include <cstdio> #include <cstring> #include <algori ...
linux下使用svn创建版本库和权限管理
linux上的svn服务端如何和本地的电脑客户端结合使用 Linux上安装SVN服务器: 第一步:检查是否已安装 # rpm -qa subversion 第二步: 通过yum命令安装svnserve ...
493 Reverse Pairs 翻转对
给定一个数组 nums ,如果 i < j 且 nums[i] > 2*nums[j] 我们就将 (i, j) 称作一个重要翻转对.你需要返回给定数组中的重要翻转对的数量.示例 1:输入: ...
Python读取文件行数不对
对于一个大文件,读取每一个行然后处理,用readline()方法老是读不全,会读到一半就结束,也不报错: 总之处理的行数跟 wc -l 统计的不一样,调试了一下午,改用 with open('xxx. ...
Java GUI 基础组件
1.JLabel 标签构造函数: JLabel() JLabel(String text) JLabel(String text,int align) //第二个参数设置文本的对齐方式,常 ...
sass 常用用法笔记
最近公司开发的h5项目,需要用到sass,所以领导推荐让我去阮一峰大神的SASS用法指南博客学习,为方便以后自己使用,所以在此记录. 一.代码的重用 1.继承:SASS允许一个选择器,继承另一个选择器 ...
洛谷 P1163 银行贷款
题目描述当一个人从银行贷款后,在一段时间内他(她)将不得不每月偿还固定的分期付款.这个问题要求计算出贷款者向银行支付的利率.假设利率按月累计. 输入输出格式输入格式: 输入文件仅一行包含三个用空格 ...

*lucene索引_的删除和更新

【删除】

【恢复删除】

【强制删除】

【优化和合并】

【更新索引】

*lucene索引_的删除和更新的更多相关文章

随机推荐

热门专题