MX2008000176A - Procesamiento de errores de colocacion en documentos. - Google Patents

Procesamiento de errores de colocacion en documentos.

Info

Publication number
MX2008000176A
MX2008000176A MX2008000176A MX2008000176A MX2008000176A MX 2008000176 A MX2008000176 A MX 2008000176A MX 2008000176 A MX2008000176 A MX 2008000176A MX 2008000176 A MX2008000176 A MX 2008000176A MX 2008000176 A MX2008000176 A MX 2008000176A
Authority
MX
Mexico
Prior art keywords
documents
sentence
query
mistakes
processing
Prior art date
Application number
MX2008000176A
Other languages
English (en)
Inventor
Hsiao-Wuen Hon
Jianfeng Gao
Ming Zhou
Original Assignee
Microsoft Corp
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Microsoft Corp filed Critical Microsoft Corp
Publication of MX2008000176A publication Critical patent/MX2008000176A/es

Links

Classifications

    • GPHYSICS
    • G06COMPUTING; CALCULATING OR COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F40/00Handling natural language data
    • G06F40/20Natural language analysis
    • G06F40/253Grammatical analysis; Style critique
    • YGENERAL TAGGING OF NEW TECHNOLOGICAL DEVELOPMENTS; GENERAL TAGGING OF CROSS-SECTIONAL TECHNOLOGIES SPANNING OVER SEVERAL SECTIONS OF THE IPC; TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
    • Y10TECHNICAL SUBJECTS COVERED BY FORMER USPC
    • Y10STECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
    • Y10S707/00Data processing: database and file management or data structures
    • Y10S707/99931Database or file accessing
    • Y10S707/99933Query processing, i.e. searching
    • YGENERAL TAGGING OF NEW TECHNOLOGICAL DEVELOPMENTS; GENERAL TAGGING OF CROSS-SECTIONAL TECHNOLOGIES SPANNING OVER SEVERAL SECTIONS OF THE IPC; TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
    • Y10TECHNICAL SUBJECTS COVERED BY FORMER USPC
    • Y10STECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
    • Y10S707/00Data processing: database and file management or data structures
    • Y10S707/99931Database or file accessing
    • Y10S707/99933Query processing, i.e. searching
    • Y10S707/99936Pattern matching access

Landscapes

  • Engineering & Computer Science (AREA)
  • Theoretical Computer Science (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Health & Medical Sciences (AREA)
  • Computational Linguistics (AREA)
  • General Health & Medical Sciences (AREA)
  • Physics & Mathematics (AREA)
  • General Engineering & Computer Science (AREA)
  • General Physics & Mathematics (AREA)
  • Artificial Intelligence (AREA)
  • Information Retrieval, Db Structures And Fs Structures Therefor (AREA)
  • Machine Translation (AREA)
  • Document Processing Apparatus (AREA)

Abstract

Una oracion es accesada y por lo menos una pregunta es generada con base en la oracion. Por lo menos una pregunta puede ser comparada con el texto dentro de una coleccion de documentos, por ejemplo, utilizando una maquina de busqueda de red. Los errores de colocacion en la oracion pueden ser detectados y/o corregidos basandose en la comparacion de por lo menos uno de la pregunta y el texto dentro de la coleccion de documentos.
MX2008000176A 2005-07-08 2006-06-30 Procesamiento de errores de colocacion en documentos. MX2008000176A (es)

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
US11/177,136 US7574348B2 (en) 2005-07-08 2005-07-08 Processing collocation mistakes in documents
PCT/US2006/026012 WO2007008492A2 (en) 2005-07-08 2006-06-30 Processing collocation mistakes in documents

Publications (1)

Publication Number Publication Date
MX2008000176A true MX2008000176A (es) 2008-04-02

Family

ID=37619276

Family Applications (1)

Application Number Title Priority Date Filing Date
MX2008000176A MX2008000176A (es) 2005-07-08 2006-06-30 Procesamiento de errores de colocacion en documentos.

Country Status (10)

Country Link
US (1) US7574348B2 (es)
EP (1) EP1899835B1 (es)
JP (1) JP5362353B2 (es)
KR (1) KR20080023341A (es)
CN (1) CN101218573A (es)
AU (1) AU2006269494A1 (es)
CA (1) CA2614416C (es)
MX (1) MX2008000176A (es)
NO (1) NO20080112L (es)
WO (1) WO2007008492A2 (es)

Families Citing this family (28)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
KR100837750B1 (ko) * 2006-08-25 2008-06-13 엔에이치엔(주) 성조를 이용하여 중국어를 검색하는 방법 및 상기 방법을수행하는 시스템
US7774193B2 (en) * 2006-12-05 2010-08-10 Microsoft Corporation Proofing of word collocation errors based on a comparison with collocations in a corpus
US20110055209A1 (en) * 2007-02-23 2011-03-03 Anthony Novac System and method for delivering content and advertisments
KR100978581B1 (ko) * 2008-05-08 2010-08-27 엔에이치엔(주) 웹 페이지 열람 중에 편리하게 사전 서비스를 제공하기위한 방법 및 시스템
US8473278B2 (en) * 2008-07-24 2013-06-25 Educational Testing Service Systems and methods for identifying collocation errors in text
US20100082324A1 (en) * 2008-09-30 2010-04-01 Microsoft Corporation Replacing terms in machine translation
US8484014B2 (en) * 2008-11-03 2013-07-09 Microsoft Corporation Retrieval using a generalized sentence collocation
TWI403911B (zh) * 2008-11-28 2013-08-01 Inst Information Industry 中文辭典建置裝置和方法,以及儲存媒體
US8250072B2 (en) * 2009-03-06 2012-08-21 Dmitri Asonov Detecting real word typos
CN101930594B (zh) * 2010-04-14 2012-05-23 山东山大鸥玛软件有限公司 一种扫描文档图像的快速纠偏方法
US20110271232A1 (en) 2010-04-30 2011-11-03 Orbis Technologies, Inc. Systems and methods for semantic search, content correlation and visualization
US10496714B2 (en) * 2010-08-06 2019-12-03 Google Llc State-dependent query response
US9262397B2 (en) 2010-10-08 2016-02-16 Microsoft Technology Licensing, Llc General purpose correction of grammatical and word usage errors
US8855997B2 (en) 2011-07-28 2014-10-07 Microsoft Corporation Linguistic error detection
US9015080B2 (en) 2012-03-16 2015-04-21 Orbis Technologies, Inc. Systems and methods for semantic inference and reasoning
US8484017B1 (en) 2012-09-10 2013-07-09 Google Inc. Identifying media content
US20140074466A1 (en) 2012-09-10 2014-03-13 Google Inc. Answering questions using environmental context
US9189531B2 (en) 2012-11-30 2015-11-17 Orbis Technologies, Inc. Ontology harmonization and mediation systems and methods
CN103365838B (zh) * 2013-07-24 2016-04-20 桂林电子科技大学 基于多元特征的英语作文语法错误自动纠正方法
US9298695B2 (en) 2013-09-05 2016-03-29 At&T Intellectual Property I, Lp Method and apparatus for managing auto-correction in messaging
CN103678714B (zh) * 2013-12-31 2017-05-10 北京百度网讯科技有限公司 实体知识库的构建方法和装置
US20160087929A1 (en) * 2014-09-24 2016-03-24 Zoho Corporation Private Limited Methods and apparatus for document creation via email
US10691709B2 (en) 2015-10-28 2020-06-23 Open Text Sa Ulc System and method for subset searching and associated search operators
US10747815B2 (en) 2017-05-11 2020-08-18 Open Text Sa Ulc System and method for searching chains of regions and associated search operators
US10241716B2 (en) 2017-06-30 2019-03-26 Microsoft Technology Licensing, Llc Global occupancy aggregator for global garbage collection scheduling
EP3649566A4 (en) 2017-07-06 2021-04-14 Open Text SA ULC SYSTEM AND PROCEDURE FOR VALUE-BASED AREA SEARCH AND RELATED SEARCH OPERATORS
US10824686B2 (en) * 2018-03-05 2020-11-03 Open Text Sa Ulc System and method for searching based on text blocks and associated search operators
US11551006B2 (en) * 2019-09-09 2023-01-10 International Business Machines Corporation Removal of personality signatures

Family Cites Families (34)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPH083815B2 (ja) * 1985-10-25 1996-01-17 株式会社日立製作所 自然言語の共起関係辞書保守方法
GB8625468D0 (en) 1986-10-24 1987-04-15 Smiths Industries Plc Speech recognition apparatus
US4868750A (en) * 1987-10-07 1989-09-19 Houghton Mifflin Company Collocational grammar system
US5251129A (en) * 1990-08-21 1993-10-05 General Electric Company Method for automated morphological analysis of word structure
US5541836A (en) * 1991-12-30 1996-07-30 At&T Corp. Word disambiguation apparatus and methods
US5383120A (en) * 1992-03-02 1995-01-17 General Electric Company Method for tagging collocations in text
US5617488A (en) * 1995-02-01 1997-04-01 The Research Foundation Of State University Of New York Relaxation word recognizer
US5887120A (en) * 1995-05-31 1999-03-23 Oracle Corporation Method and apparatus for determining theme for discourse
US5680511A (en) * 1995-06-07 1997-10-21 Dragon Systems, Inc. Systems and methods for word recognition
US5721938A (en) 1995-06-07 1998-02-24 Stuckey; Barbara K. Method and device for parsing and analyzing natural language sentences and text
US5907839A (en) * 1996-07-03 1999-05-25 Yeda Reseach And Development, Co., Ltd. Algorithm for context sensitive spelling correction
US6173298B1 (en) * 1996-09-17 2001-01-09 Asap, Ltd. Method and apparatus for implementing a dynamic collocation dictionary
CN1193779A (zh) * 1997-03-13 1998-09-23 国际商业机器公司 中文语句分词方法及其在中文查错系统中的应用
GB2329047A (en) 1997-09-05 1999-03-10 Sharp Kk A method of identifying collocates
KR980004126A (ko) 1997-12-16 1998-03-30 양승택 다국어 웹 문서 검색을 위한 질의어 변환 장치 및 방법
GB2334115A (en) 1998-01-30 1999-08-11 Sharp Kk Processing text eg for approximate translation
US6216123B1 (en) * 1998-06-24 2001-04-10 Novell, Inc. Method and system for rapid retrieval in a full text indexing system
GB9821787D0 (en) * 1998-10-06 1998-12-02 Data Limited Apparatus for classifying or processing data
JP2001101186A (ja) * 1999-09-30 2001-04-13 Oki Electric Ind Co Ltd 機械翻訳装置
GB0006721D0 (en) 2000-03-20 2000-05-10 Mitchell Thomas A Assessment methods and systems
US7860706B2 (en) * 2001-03-16 2010-12-28 Eli Abir Knowledge system method and appparatus
US20020152219A1 (en) 2001-04-16 2002-10-17 Singh Monmohan L. Data interexchange protocol
US7269546B2 (en) * 2001-05-09 2007-09-11 International Business Machines Corporation System and method of finding documents related to other documents and of finding related words in response to a query to refine a search
US7003444B2 (en) * 2001-07-12 2006-02-21 Microsoft Corporation Method and apparatus for improved grammar checking using a stochastic parser
US7246060B2 (en) * 2001-11-06 2007-07-17 Microsoft Corporation Natural input recognition system and method using a contextual mapping engine and adaptive user bias
US20030154071A1 (en) 2002-02-11 2003-08-14 Shreve Gregory M. Process for the document management and computer-assisted translation of documents utilizing document corpora constructed by intelligent agents
KR100530154B1 (ko) 2002-06-07 2005-11-21 인터내셔널 비지네스 머신즈 코포레이션 변환방식 기계번역시스템에서 사용되는 변환사전을생성하는 방법 및 장치
US7031911B2 (en) 2002-06-28 2006-04-18 Microsoft Corporation System and method for automatic detection of collocation mistakes in documents
US7171351B2 (en) * 2002-09-19 2007-01-30 Microsoft Corporation Method and system for retrieving hint sentences using expanded queries
US7249012B2 (en) * 2002-11-20 2007-07-24 Microsoft Corporation Statistical method and apparatus for learning translation relationships among phrases
US7689412B2 (en) 2003-12-05 2010-03-30 Microsoft Corporation Synonymous collocation extraction using translation information
US7707039B2 (en) * 2004-02-15 2010-04-27 Exbiblio B.V. Automatic modification of web pages
US20060282255A1 (en) * 2005-06-14 2006-12-14 Microsoft Corporation Collocation translation from monolingual and available bilingual corpora
US20070016397A1 (en) * 2005-07-18 2007-01-18 Microsoft Corporation Collocation translation using monolingual corpora

Also Published As

Publication number Publication date
EP1899835B1 (en) 2019-06-26
AU2006269494A1 (en) 2007-01-18
CA2614416A1 (en) 2007-01-18
JP5362353B2 (ja) 2013-12-11
EP1899835A2 (en) 2008-03-19
EP1899835A4 (en) 2017-10-25
CA2614416C (en) 2014-05-27
JP2009500754A (ja) 2009-01-08
WO2007008492A2 (en) 2007-01-18
WO2007008492A3 (en) 2007-06-21
CN101218573A (zh) 2008-07-09
NO20080112L (no) 2008-02-01
US20070010992A1 (en) 2007-01-11
US7574348B2 (en) 2009-08-11
KR20080023341A (ko) 2008-03-13

Similar Documents

Publication Publication Date Title
MX2008000176A (es) Procesamiento de errores de colocacion en documentos.
WO2007108788A3 (en) Method and system for answer extraction
WO2007035912A3 (en) Document processing
WO2012048306A3 (en) Structured searching of dynamic structured document corpuses
WO2006116196A3 (en) Media object metadata association and ranking
WO2008051750A3 (en) Associating geographic-related information with objects
WO2007137145A3 (en) Certificate-based search
MX2010002349A (es) Resolucion de correferencia en un sistema de procesamiento de lenguaje natural sensible a la ambiguedad.
WO2006110684A3 (en) System and method for searching for a query
WO2008100849A3 (en) Semantics-based method and system for document analysis
WO2008156473A3 (en) Using relevance feedback in face recognition
WO2012047579A3 (en) Representing and processing inter-slot constraints on component selection for dynamic ads
WO2006132793A3 (en) Learning facts from semi-structured text
WO2007019311A3 (en) Systems for and methods of finding relevant documents by analyzing tags
WO2010120929A3 (en) Generating user-customized search results and building a semantics-enhanced search engine
WO2011019749A3 (en) Presenting comments from various sources
WO2012058690A3 (en) Transforming search engine queries
WO2007130544A3 (en) Method for domain identification of documents in a document database
WO2009003072A3 (en) Integrated platform for user input of digital ink
WO2008130952A3 (en) Extensible database system and method
NZ598238A (en) Image element searching
ATE418106T1 (de) Generisches suchverfahren für verschiedenen objekttypen
WO2006110373A3 (en) Apparatus and method for utilizing sentence component metadata to create database queries
WO2011011777A3 (en) Pre-computed ranking using proximity terms
Berglund et al. Internmedicin

Legal Events

Date Code Title Description
FA Abandonment or withdrawal