Java OriginalTextAnnotation类代码示例

OStack程序员社区-中国程序员成长平台 › 门户 › 编程› Java›Java编程经验

原作者: [db:作者] 来自: [db:来源] 收藏邀请

本文整理汇总了Java中edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation类的典型用法代码示例。如果您正苦于以下问题：Java OriginalTextAnnotation类的具体用法？Java OriginalTextAnnotation怎么用？Java OriginalTextAnnotation使用的例子？那么恭喜您, 这里精选的类代码示例或许可以为您提供帮助。

OriginalTextAnnotation类属于edu.stanford.nlp.ling.CoreAnnotations包，在下文中一共展示了OriginalTextAnnotation类的10个代码示例，这些例子默认根据受欢迎程度排序。您可以为喜欢或者感觉有用的代码点赞，您的评价将有助于我们的系统推荐出更棒的Java代码示例。

示例1: TokenizedCoreLabelWrapper

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
/**
 *
 */
public TokenizedCoreLabelWrapper(final CoreLabel cl) {
  this.value = cl.get(ValueAnnotation.class);
  this.text = cl.get(TextAnnotation.class);
  LOGGER.trace("Wrapping token text: {}", this.text);
  this.originalText = cl.get(OriginalTextAnnotation.class);
  this.before = cl.get(BeforeAnnotation.class);
  this.after = cl.get(AfterAnnotation.class);

  this.startSentenceOffset = cl.get(CharacterOffsetBeginAnnotation.class);
  this.endSentenceOffset = cl.get(CharacterOffsetEndAnnotation.class);

  this.startOffset = Optional.ofNullable(cl.get(TokenBeginAnnotation.class));
  this.endOffset = Optional.ofNullable(cl.get(TokenEndAnnotation.class));
  LOGGER.trace("TokenBegin: {}", this.startOffset);
  LOGGER.trace("TokenEnd: {}", this.endOffset);

  this.idx = cl.get(IndexAnnotation.class);
  this.sentenceIdx = cl.get(SentenceIndexAnnotation.class);
  LOGGER.trace("Got sentence idx: {}", this.sentenceIdx);
}

开发者ID:hltcoe，项目名称:concrete-stanford-deprecated2，代码行数:24，代码来源:TokenizedCoreLabelWrapper.java

示例2: getNext

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
/** Make the next token.
 *
 *  @param txt What the token should be
 *  @param originalText The original String that got transformed into txt
 */
private Object getNext(String txt, String originalText) {
  if (tokenFactory == null) {
    throw new RuntimeException(this.getClass().getName() + ": Token factory is null.");
  }
  if (invertible) {
    //String str = prevWordAfter.toString();
    //prevWordAfter.setLength(0);
    CoreLabel word = (CoreLabel) tokenFactory.makeToken(txt, yychar, yylength());
    word.set(OriginalTextAnnotation.class, originalText);
    //word.set(BeforeAnnotation.class, str);
    //prevWord.set(AfterAnnotation.class, str);
    //prevWord = word;
    return word;
  } else {
    return tokenFactory.makeToken(txt, yychar, yylength());
  }
}

开发者ID:amark-india，项目名称:eventspotter，代码行数:23，代码来源:ArabicLexer.java

示例3: buildMention

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
public Mention buildMention(Annotation annotation, int sentId,
		int startToken, int endToken) {
	CoreMap sentAnn = annotation.get(SentencesAnnotation.class).get(sentId);
	List<CoreLabel> tokens = sentAnn.get(TokensAnnotation.class);
	// create a Mention object
	Mention.Builder m = Mention.newBuilder();
	m.setStart(startToken);
	m.setEnd(endToken);
	for (int i = 0; i < tokens.size(); i++) {
		m.addTokens(tokens.get(i).get(OriginalTextAnnotation.class));
		m.addPosTags(tokens.get(i).get(PartOfSpeechAnnotation.class));
	}
	m.setEntityName("");
	m.setFileid("on-the-fly");
	m.setSentid(sentId);

	// dependency
	String depStr = StanfordDependencyResolver.getString(sentAnn);
	if (depStr != null) {
		for (String d : depStr.split("\t")) {
			Matcher match = Preprocessing.depPattern.matcher(d);
			if (match.find()) {
				m.addDeps(Dependency.newBuilder().setType(match.group(1))
						.setGov(Integer.parseInt(match.group(3)) - 1)
						.setDep(Integer.parseInt(match.group(5)) - 1)
						.build());
			} else {

			}
		}
	}
	return m.build();
}

开发者ID:zhangcongle，项目名称:NewsSpikeRe，代码行数:34，代码来源:FigerSystem.java

示例4: buildMention

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
public Mention buildMention(Annotation annotation, int sentId, int startToken, int endToken) {
	CoreMap sentAnn = annotation.get(SentencesAnnotation.class).get(sentId);
	List<CoreLabel> tokens = sentAnn.get(TokensAnnotation.class);
	// create a Mention object
	Mention.Builder m = Mention.newBuilder();
	m.setStart(startToken);
	m.setEnd(endToken);
	for (int i = 0; i < tokens.size(); i++) {
		m.addTokens(tokens.get(i).get(OriginalTextAnnotation.class));
		m.addPosTags(tokens.get(i).get(PartOfSpeechAnnotation.class));
	}
	m.setEntityName("");
	m.setFileid("on-the-fly");
	m.setSentid(sentId);

	// dependency
	String depStr = StanfordDependencyResolver.getString(sentAnn);
	if (depStr != null) {
		for (String d : depStr.split("\t")) {
			Matcher match = Preprocessing.depPattern.matcher(d);
			if (match.find()) {
				m.addDeps(
						Dependency.newBuilder().setType(match.group(1)).setGov(Integer.parseInt(match.group(3)) - 1)
								.setDep(Integer.parseInt(match.group(5)) - 1).build());
			} else {

			}
		}
	}
	return m.build();
}

开发者ID:zhangcongle，项目名称:NewsSpikeRe，代码行数:32，代码来源:ParseStanfordFigerReverb.java

示例5: process

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
@Override
public SymmetricalWordAlignment process(Sequence<IString> inputSequence) {
  List<CoreLabel> labeledTokens = ProcessorTools.toCharacterSequence(inputSequence);
  labeledTokens = classifier.classify(labeledTokens);
  List<CoreLabel> outputTokens = ProcessorTools.toPostProcessedSequence(labeledTokens);
  List<String> outputStrings = new ArrayList<>();
  for (CoreLabel label : outputTokens) {
    outputStrings.add(label.word());
  }
  Sequence<IString> outputSequence = IStrings.toIStringSequence(outputStrings);
  SymmetricalWordAlignment alignment = new SymmetricalWordAlignment(inputSequence, outputSequence);
  
  // Reconstruct the alignment by iterating over the target
  for (int outputIndex = 0, inputIndex = 0, outputSize = outputTokens.size(), inputSize = inputSequence.size(); 
      outputIndex < outputSize && inputIndex < inputSize; ++outputIndex) {
    String outputToken = outputTokens.get(outputIndex).get(OriginalTextAnnotation.class);
    String inputCandidate = inputSequence.get(inputIndex).toString();
    if (outputToken.equals(inputCandidate)) {
      // Unigram alignment
      alignment.addAlign(inputIndex, outputIndex);
      ++inputIndex;
      
    } else {
      // one-to-many alignment (output-to-input)
      String[] originalTokens = outputToken.split("\\s+");
      int newInputIndex = findSubsequenceStart(originalTokens, inputSequence, inputIndex);
      if (newInputIndex < 0) {
        System.err.printf("Unable to find |%s| in |%s|%n", outputToken, inputSequence.toString());
      } else {
        inputIndex = newInputIndex;
        for (int i = 0; i < originalTokens.length; ++i) {
          alignment.addAlign(inputIndex, outputIndex);
          ++inputIndex;
        }
      }
    }
  }
  return alignment;
}

开发者ID:stanfordnlp，项目名称:phrasal，代码行数:40，代码来源:CRFPostprocessor.java

示例6: getNext

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
/**
 * Make the next token.
 * 
 * @param txt
 *          What the token should be
 * @param current
 *          The original String that got transformed into txt
 */
private Object getNext(String txt, String current) {
  if (invertible) {
    CoreLabel word = (CoreLabel) tokenFactory.makeToken(txt, yychar, yylength());
    word.set(OriginalTextAnnotation.class, current);
    word.set(BeforeAnnotation.class, prevWord.getString(AfterAnnotation.class));
    prevWord = word;
    return word;
  } else {
    return tokenFactory.makeToken(txt, yychar, yylength());
  }
}

开发者ID:begab，项目名称:kpe，代码行数:20，代码来源:HunPTBLexer.java

示例7: NGram

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
public NGram(String[] originalForms, String[] lemmas) {
  for (int i = 0; i < originalForms.length; ++i) {
    CoreLabel cl = new CoreLabel();
    cl.set(TextAnnotation.class, originalForms[i]);
    cl.set(OriginalTextAnnotation.class, originalForms[i]);
    cl.set(LemmaAnnotation.class, lemmas[i]);
    cl.set(PartOfSpeechAnnotation.class, "dummy");
    add(cl);
  }
  setNormalizedForm();
}

开发者ID:begab，项目名称:kpe，代码行数:12，代码来源:NGram.java

示例8: getNext

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
/** Make the next token.
 *  @param txt What the token should be
 *  @param originalText The original String that got transformed into txt
 */
private Object getNext(String txt, String originalText) {
  if (invertible) {
    String str = prevWordAfter.toString();
    prevWordAfter.setLength(0);
    CoreLabel word = (CoreLabel) tokenFactory.makeToken(txt, yychar, yylength());
    word.set(OriginalTextAnnotation.class, originalText);
    word.set(BeforeAnnotation.class, str);
    prevWord.set(AfterAnnotation.class, str);
    prevWord = word;
    return word;
  } else {
    return tokenFactory.makeToken(txt, yychar, yylength());
 }
}

开发者ID:amark-india，项目名称:eventspotter，代码行数:19，代码来源:PTBLexer.java

示例9: setOriginalText

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
/**
 * {@inheritDoc}
 */
public void setOriginalText(String originalText) {
  set(CoreAnnotations.OriginalTextAnnotation.class, originalText);
}

开发者ID:amark-india，项目名称:eventspotter，代码行数:7，代码来源:CoreLabel.java

示例10: originalText

import edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation; //导入依赖的package包/类
/**
 * {@inheritDoc}
 */
public String originalText() {
  return getString(OriginalTextAnnotation.class);
}

开发者ID:amark-india，项目名称:eventspotter，代码行数:7，代码来源:CoreLabel.java

注：本文中的edu.stanford.nlp.ling.CoreAnnotations.OriginalTextAnnotation类示例整理自Github/MSDocs等源码及文档管理平台，相关代码片段筛选自各路编程大神贡献的开源项目，源码版权归原作者所有，传播和使用请参考对应项目的License；未经允许，请勿转载。

鲜花

握手

雷人

路过

鸡蛋

该文章已有0人参与评论

请发表评论

全部评论

专题导读

More+

10-27 六六分期app的软件客服如何联系？(六六分期

11-06 可心卡盟:win10系统火狐flash插件崩溃怎么

11-06 亲亲特价:怎么删除回收站图标

11-06 济南大学虚拟社区:鲁大师节能降温的具体办

11-06 xlueops.exe:无线网络安装向导

11-06 女斗合众国:win7系统cf与主机连接不稳定怎

11-06 0xc000022-[cf烟雾头]cf怎么调烟雾头

11-06 qizideyouhuo:应用程序无法正常启动0xc0000

11-06 ipz-185:win7系统vcf文件怎么打开

11-06 傻哥蹦迪:win10系统s4怎么打开usb调试

11-06 八神浩树gtaste:回收站清空了怎么恢复

11-06 妖尾之黑色守护:win10系统电脑没有1440x900

11-06 校园至尊魔王小说:win7系统浏览网页时字体

11-06 女斗合众国:win10系统访问共享文件夹提示请

11-06 tokyo hot n0654:恢复win7系统默认字体一招

11-06 雨酷仙境:设置win7系统转移临时文件夹腾出

11-06 阿穆纳伊之杖:win7系统开始菜单在右边还原

11-06 tunespotting:win10系统火狐flash插件总是

11-06 甘尔葛分析师：计谋网站seo关键词暴涨有什

11-06 蔡贵霖: 计谋网站seo关键词暴涨有什么秘密

11-06 博益网首页:ao3网页版进入不了解决方法

11-06 漏斗子专栏: 网站数据分析小白易懂精华篇

11-06 见证双虹怎么做:win7系统开启telnet命令的

11-06 颾狐蝶蜋:系统资源不足无法完成请求的服务

11-06 国光中学校歌:提交网站到alexa查询详细步骤

11-06 西安有情天:静态网页和动态网页的区别

11-06 红木雅尚斋:外部链接构造对网站的好处

11-06 前官礼遇：防止域名劫持–增强域安全性的10

11-06 密传二转答案: 中文分词算法有哪些

11-06 金泉家园邮编:百度快照劫持的表现及应对方

Java TrueFilter类代码示例发布时间：2022-05-22

Java RegisterSpecList类代码示例发布时间：2022-05-22

剪的笔顺,诠释剪的笔画,认识剪的部首

1 六六分期app的软件客服如何联系？(六六分期

六六分期app的软件客服如何联系？不知道吗？加qq群【895510560】即可！标题：六六分期

阅读：19212|2023-10-27

2 可心卡盟:win10系统火狐flash插件崩溃怎么

今天小编告诉大家如何处理win10系统火狐flash插件总是崩溃的问题，可能很多用户都不知

阅读：9994|2022-11-06

3 亲亲特价:怎么删除回收站图标

今天小编告诉大家如何对win10系统删除桌面回收站图标进行设置，可能很多用户都不知道

阅读：8331|2022-11-06

4 济南大学虚拟社区:鲁大师节能降温的具体办

今天小编告诉大家如何对win10系统电脑设置节能降温的设置方法，想必大家都遇到过需要

阅读：8699|2022-11-06

5 xlueops.exe:无线网络安装向导

我们在使用xp系统的过程中,经常需要对xp系统无线网络安装向导设置进行设置，可能很多

阅读：8642|2022-11-06

6 女斗合众国:win7系统cf与主机连接不稳定怎

今天小编告诉大家如何处理win7系统玩cf老是与主机连接不稳定的问题，可能很多用户都不

阅读：9662|2022-11-06

7 0xc000022-[cf烟雾头]cf怎么调烟雾头

电脑对日常生活的重要性小编就不多说了，可是一旦碰到win7系统设置cf烟雾头的问题，很

阅读：8627|2022-11-06

8 qizideyouhuo:应用程序无法正常启动0xc0000

我们在日常使用电脑的时候，有的小伙伴们可能在打开应用的时候会遇见提示应用程序无法

阅读：8003|2022-11-06

9 ipz-185:win7系统vcf文件怎么打开

今天小编告诉大家如何对win7系统打开vcf文件进行设置，可能很多用户都不知道怎么对win

阅读：8660|2022-11-06

10 傻哥蹦迪:win10系统s4怎么打开usb调试

今天小编告诉大家如何对win10系统s4开启USB调试模式进行设置，可能很多用户都不知道怎

阅读：7537|2022-11-06

客服电话

电子邮件

Java OriginalTextAnnotation类代码示例

示例1: TokenizedCoreLabelWrapper

示例2: getNext

示例3: buildMention

示例4: buildMention

示例5: process

示例6: getNext

示例7: NGram

示例8: getNext

示例9: setOriginalText

示例10: originalText

请发表评论

全部评论

上一篇：

下一篇：

sb2nov/mac-setup: Installing Development

belane/linux-soft-exploit-suggester: Sea

harningt/luajson: JSON parser/encoder fo

mvallieres/radiomics: MATLAB programming

Delphi 的RTTI机制浅探

剪的笔顺,诠释剪的笔画,认识剪的部首

六六分期app的软件客服如何联系？(六六分期

florent37/ViewAnimator: A fluent Android

florent37/Shrine-MaterialDesign2: implem

CVE-2020-36276

SimpleSoftwareIO/simple-sms: Send and re

关于我们

产品与服务

解决方案

139-2527-9053