NO20080112L - Behandling av feil i samlokaliseringen i dokumenter - Google Patents

Behandling av feil i samlokaliseringen i dokumenter

Info

Publication number
NO20080112L
NO20080112L NO20080112A NO20080112A NO20080112L NO 20080112 L NO20080112 L NO 20080112L NO 20080112 A NO20080112 A NO 20080112A NO 20080112 A NO20080112 A NO 20080112A NO 20080112 L NO20080112 L NO 20080112L
Authority
NO
Norway
Prior art keywords
documents
location
processing errors
query
statement
Prior art date
Application number
NO20080112A
Other languages
English (en)
Inventor
Hsiao-Wuen Hon
Jianfeng Gao
Ming Zhou
Original Assignee
Microsoft Corp
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Microsoft Corp filed Critical Microsoft Corp
Publication of NO20080112L publication Critical patent/NO20080112L/no

Links

Classifications

    • GPHYSICS
    • G06COMPUTING; CALCULATING OR COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F40/00Handling natural language data
    • G06F40/20Natural language analysis
    • G06F40/253Grammatical analysis; Style critique
    • YGENERAL TAGGING OF NEW TECHNOLOGICAL DEVELOPMENTS; GENERAL TAGGING OF CROSS-SECTIONAL TECHNOLOGIES SPANNING OVER SEVERAL SECTIONS OF THE IPC; TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
    • Y10TECHNICAL SUBJECTS COVERED BY FORMER USPC
    • Y10STECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
    • Y10S707/00Data processing: database and file management or data structures
    • Y10S707/99931Database or file accessing
    • Y10S707/99933Query processing, i.e. searching
    • YGENERAL TAGGING OF NEW TECHNOLOGICAL DEVELOPMENTS; GENERAL TAGGING OF CROSS-SECTIONAL TECHNOLOGIES SPANNING OVER SEVERAL SECTIONS OF THE IPC; TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
    • Y10TECHNICAL SUBJECTS COVERED BY FORMER USPC
    • Y10STECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
    • Y10S707/00Data processing: database and file management or data structures
    • Y10S707/99931Database or file accessing
    • Y10S707/99933Query processing, i.e. searching
    • Y10S707/99936Pattern matching access

Landscapes

  • Engineering & Computer Science (AREA)
  • Theoretical Computer Science (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Health & Medical Sciences (AREA)
  • Computational Linguistics (AREA)
  • General Health & Medical Sciences (AREA)
  • Physics & Mathematics (AREA)
  • General Engineering & Computer Science (AREA)
  • General Physics & Mathematics (AREA)
  • Artificial Intelligence (AREA)
  • Information Retrieval, Db Structures And Fs Structures Therefor (AREA)
  • Machine Translation (AREA)
  • Document Processing Apparatus (AREA)

Abstract

En setning leses og minst én spørring blir generert basert på setningen. Minst én spørring kan bli sammenliknet med tekst fra en samling av dokumenter, for eksempel ved anvendelse av en webbasert søkemotor. Kollokasjonsfeil i setningen kan oppdages og/eller rettes opp basert på sammenlikningen av den minst ene spørringen og teksten fra samlingen av dokumenter.
NO20080112A 2005-07-08 2008-01-08 Behandling av feil i samlokaliseringen i dokumenter NO20080112L (no)

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
US11/177,136 US7574348B2 (en) 2005-07-08 2005-07-08 Processing collocation mistakes in documents
PCT/US2006/026012 WO2007008492A2 (en) 2005-07-08 2006-06-30 Processing collocation mistakes in documents

Publications (1)

Publication Number Publication Date
NO20080112L true NO20080112L (no) 2008-02-01

Family

ID=37619276

Family Applications (1)

Application Number Title Priority Date Filing Date
NO20080112A NO20080112L (no) 2005-07-08 2008-01-08 Behandling av feil i samlokaliseringen i dokumenter

Country Status (10)

Country Link
US (1) US7574348B2 (no)
EP (1) EP1899835B1 (no)
JP (1) JP5362353B2 (no)
KR (1) KR20080023341A (no)
CN (1) CN101218573A (no)
AU (1) AU2006269494A1 (no)
CA (1) CA2614416C (no)
MX (1) MX2008000176A (no)
NO (1) NO20080112L (no)
WO (1) WO2007008492A2 (no)

Families Citing this family (28)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
KR100837750B1 (ko) * 2006-08-25 2008-06-13 엔에이치엔(주) 성조를 이용하여 중국어를 검색하는 방법 및 상기 방법을수행하는 시스템
US7774193B2 (en) * 2006-12-05 2010-08-10 Microsoft Corporation Proofing of word collocation errors based on a comparison with collocations in a corpus
CA2679094A1 (en) * 2007-02-23 2008-08-28 1698413 Ontario Inc. System and method for delivering content and advertisements
KR100978581B1 (ko) * 2008-05-08 2010-08-27 엔에이치엔(주) 웹 페이지 열람 중에 편리하게 사전 서비스를 제공하기위한 방법 및 시스템
US8473278B2 (en) * 2008-07-24 2013-06-25 Educational Testing Service Systems and methods for identifying collocation errors in text
US20100082324A1 (en) * 2008-09-30 2010-04-01 Microsoft Corporation Replacing terms in machine translation
US8484014B2 (en) * 2008-11-03 2013-07-09 Microsoft Corporation Retrieval using a generalized sentence collocation
TWI403911B (zh) * 2008-11-28 2013-08-01 Inst Information Industry 中文辭典建置裝置和方法,以及儲存媒體
US8250072B2 (en) * 2009-03-06 2012-08-21 Dmitri Asonov Detecting real word typos
CN101930594B (zh) * 2010-04-14 2012-05-23 山东山大鸥玛软件有限公司 一种扫描文档图像的快速纠偏方法
US8725771B2 (en) * 2010-04-30 2014-05-13 Orbis Technologies, Inc. Systems and methods for semantic search, content correlation and visualization
US10496714B2 (en) * 2010-08-06 2019-12-03 Google Llc State-dependent query response
US9262397B2 (en) 2010-10-08 2016-02-16 Microsoft Technology Licensing, Llc General purpose correction of grammatical and word usage errors
US8855997B2 (en) 2011-07-28 2014-10-07 Microsoft Corporation Linguistic error detection
US9015080B2 (en) 2012-03-16 2015-04-21 Orbis Technologies, Inc. Systems and methods for semantic inference and reasoning
US8484017B1 (en) 2012-09-10 2013-07-09 Google Inc. Identifying media content
US20140074466A1 (en) 2012-09-10 2014-03-13 Google Inc. Answering questions using environmental context
US9189531B2 (en) 2012-11-30 2015-11-17 Orbis Technologies, Inc. Ontology harmonization and mediation systems and methods
CN103365838B (zh) * 2013-07-24 2016-04-20 桂林电子科技大学 基于多元特征的英语作文语法错误自动纠正方法
US9298695B2 (en) 2013-09-05 2016-03-29 At&T Intellectual Property I, Lp Method and apparatus for managing auto-correction in messaging
CN103678714B (zh) * 2013-12-31 2017-05-10 北京百度网讯科技有限公司 实体知识库的构建方法和装置
US20160087929A1 (en) * 2014-09-24 2016-03-24 Zoho Corporation Private Limited Methods and apparatus for document creation via email
US10691709B2 (en) 2015-10-28 2020-06-23 Open Text Sa Ulc System and method for subset searching and associated search operators
US10747815B2 (en) 2017-05-11 2020-08-18 Open Text Sa Ulc System and method for searching chains of regions and associated search operators
US10241716B2 (en) 2017-06-30 2019-03-26 Microsoft Technology Licensing, Llc Global occupancy aggregator for global garbage collection scheduling
EP3649566A4 (en) 2017-07-06 2021-04-14 Open Text SA ULC SYSTEM AND PROCEDURE FOR VALUE-BASED AREA SEARCH AND RELATED SEARCH OPERATORS
US10824686B2 (en) 2018-03-05 2020-11-03 Open Text Sa Ulc System and method for searching based on text blocks and associated search operators
US11551006B2 (en) * 2019-09-09 2023-01-10 International Business Machines Corporation Removal of personality signatures

Family Cites Families (34)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPH083815B2 (ja) * 1985-10-25 1996-01-17 株式会社日立製作所 自然言語の共起関係辞書保守方法
GB8625468D0 (en) 1986-10-24 1987-04-15 Smiths Industries Plc Speech recognition apparatus
US4868750A (en) * 1987-10-07 1989-09-19 Houghton Mifflin Company Collocational grammar system
US5251129A (en) * 1990-08-21 1993-10-05 General Electric Company Method for automated morphological analysis of word structure
US5541836A (en) * 1991-12-30 1996-07-30 At&T Corp. Word disambiguation apparatus and methods
US5383120A (en) * 1992-03-02 1995-01-17 General Electric Company Method for tagging collocations in text
US5617488A (en) * 1995-02-01 1997-04-01 The Research Foundation Of State University Of New York Relaxation word recognizer
US5887120A (en) * 1995-05-31 1999-03-23 Oracle Corporation Method and apparatus for determining theme for discourse
US5721938A (en) * 1995-06-07 1998-02-24 Stuckey; Barbara K. Method and device for parsing and analyzing natural language sentences and text
US5680511A (en) * 1995-06-07 1997-10-21 Dragon Systems, Inc. Systems and methods for word recognition
US5907839A (en) * 1996-07-03 1999-05-25 Yeda Reseach And Development, Co., Ltd. Algorithm for context sensitive spelling correction
US6173298B1 (en) * 1996-09-17 2001-01-09 Asap, Ltd. Method and apparatus for implementing a dynamic collocation dictionary
CN1193779A (zh) * 1997-03-13 1998-09-23 国际商业机器公司 中文语句分词方法及其在中文查错系统中的应用
GB2329047A (en) * 1997-09-05 1999-03-10 Sharp Kk A method of identifying collocates
KR980004126A (ko) * 1997-12-16 1998-03-30 양승택 다국어 웹 문서 검색을 위한 질의어 변환 장치 및 방법
GB2334115A (en) * 1998-01-30 1999-08-11 Sharp Kk Processing text eg for approximate translation
US6216123B1 (en) * 1998-06-24 2001-04-10 Novell, Inc. Method and system for rapid retrieval in a full text indexing system
GB9821787D0 (en) * 1998-10-06 1998-12-02 Data Limited Apparatus for classifying or processing data
JP2001101186A (ja) * 1999-09-30 2001-04-13 Oki Electric Ind Co Ltd 機械翻訳装置
GB0006721D0 (en) * 2000-03-20 2000-05-10 Mitchell Thomas A Assessment methods and systems
US7860706B2 (en) * 2001-03-16 2010-12-28 Eli Abir Knowledge system method and appparatus
US20020152219A1 (en) * 2001-04-16 2002-10-17 Singh Monmohan L. Data interexchange protocol
US7269546B2 (en) * 2001-05-09 2007-09-11 International Business Machines Corporation System and method of finding documents related to other documents and of finding related words in response to a query to refine a search
US7003444B2 (en) * 2001-07-12 2006-02-21 Microsoft Corporation Method and apparatus for improved grammar checking using a stochastic parser
US7246060B2 (en) * 2001-11-06 2007-07-17 Microsoft Corporation Natural input recognition system and method using a contextual mapping engine and adaptive user bias
US20030154071A1 (en) * 2002-02-11 2003-08-14 Shreve Gregory M. Process for the document management and computer-assisted translation of documents utilizing document corpora constructed by intelligent agents
KR100530154B1 (ko) * 2002-06-07 2005-11-21 인터내셔널 비지네스 머신즈 코포레이션 변환방식 기계번역시스템에서 사용되는 변환사전을생성하는 방법 및 장치
US7031911B2 (en) * 2002-06-28 2006-04-18 Microsoft Corporation System and method for automatic detection of collocation mistakes in documents
US7171351B2 (en) * 2002-09-19 2007-01-30 Microsoft Corporation Method and system for retrieving hint sentences using expanded queries
US7249012B2 (en) * 2002-11-20 2007-07-24 Microsoft Corporation Statistical method and apparatus for learning translation relationships among phrases
US7689412B2 (en) * 2003-12-05 2010-03-30 Microsoft Corporation Synonymous collocation extraction using translation information
US7707039B2 (en) * 2004-02-15 2010-04-27 Exbiblio B.V. Automatic modification of web pages
US20060282255A1 (en) * 2005-06-14 2006-12-14 Microsoft Corporation Collocation translation from monolingual and available bilingual corpora
US20070016397A1 (en) * 2005-07-18 2007-01-18 Microsoft Corporation Collocation translation using monolingual corpora

Also Published As

Publication number Publication date
JP5362353B2 (ja) 2013-12-11
MX2008000176A (es) 2008-04-02
CN101218573A (zh) 2008-07-09
US7574348B2 (en) 2009-08-11
EP1899835A4 (en) 2017-10-25
KR20080023341A (ko) 2008-03-13
CA2614416C (en) 2014-05-27
AU2006269494A1 (en) 2007-01-18
EP1899835B1 (en) 2019-06-26
US20070010992A1 (en) 2007-01-11
CA2614416A1 (en) 2007-01-18
WO2007008492A3 (en) 2007-06-21
EP1899835A2 (en) 2008-03-19
JP2009500754A (ja) 2009-01-08
WO2007008492A2 (en) 2007-01-18

Similar Documents

Publication Publication Date Title
NO20080112L (no) Behandling av feil i samlokaliseringen i dokumenter
GB2473374A (en) Method of e-mail address search and e-mail address transliteration and associated device
ASGARI et al. Forensic discourse analysis: legal speech acts in legal language
NAGHZGOOY Modal verbs and the concept of exponence in Persian
Ebrahimi Comparison of Advisory Themes of Minoo-ye Kherad and Afarin Name of Abu Shakoor Balkhi
Taheri On Fronting of ū in Lori and Central Iranian Dialects
Namugambe An automated accession register for the Ministry of Health library
RASEKH et al. DEFINITENESS AND OBJECT MARKING IN TATI, TALESHI AND BALOUCHI
Koushei et al. Ḥ adī th in Islamic period art historiography: obstacles and restrictions
BOZORG et al. Necessity of Recension, Correction and Analysis of Bostan ol-Arefin va Tohfat ol-Moridin
Gandomkar Polysemy of verb “xordan”: A case study of inefficiency of lexical typology
Saeidi An investigation into the origins of a number of the stories in the book Ahval va akhbar-e Barmakian and the removal of some of its shortcomings based on secondary sources
PIRNAJMUDDIN et al. LINGUISTIC DEVIATION IN TRANSLATION OF BLANK VERSE: THE CASE OF AHMAD SHAMLU'S POETRY IN TRANSLATION
MORADI INVESTIGATION OF WORDS ENDING IN SUFFIXES «-ЯК (A)/-AК (A)» IN RUSSIAN LANGUAGE WITH COMMON GENERIC NOUN OF THIS LANGUAGE, AND COMPOUND AGENTIVE ADJECTIVE IN PERSIAN
DARZI et al. THE QUANTIFICATION PROPERTY OF THE PLURAL MARKER “-HA” IN PERSIAN
SAYYEDAN COMPARATIVE CRITICISM OF PERSIAN TRANSLATIONS OF AL-AJNIHA AL-MUTAKASSIRA GIBRAN KHALIL GIBRAN’S
MOHSENINIA et al. RESISTANCE THEME IN THE POETRY OF AHMAD AL-SAFIAL-NAJAFI (SCHOLARLY-RESEARCH)
VARASTEHFAR Are fairy tales addressed to Children?
LATIF et al. The Stages of Textual Perception in Sepehrie’s Poetry: Based on Beaugrande & Dressler’s Approach
KORD et al. PHONOLOGICAL PROCESSES IN COMPLEX AND COMPOUND WORDS
Horri TRANSLATOR’S PRESENCE IN TRANSLATED NARRATIVE TEXTS THROUGH SHIFTS PROPOSED BY LEUVEN-ZWART (3): DESCRIPTIVE MODEL
GHOLAMI et al. A STUDY OF DOMESTICATION AND FOREIGNIZATION STRATEGIES ON THE BASIS OF BERMAN'S
SEIF et al. STUDYING BALUCHI PROVERBS AND LOOKING FOR ITS ROOTS IN PERSIAN LANGUAGE
SAEIDI TRANSLATION OF SEMANTIC, PRAGMATIC AND STYLISTIC FUNCTIONS OF PUNCTUATION MARKS IN DIFFERENT TEXT-TYPE
BAYANLOU et al. THE POETICAL LICENSE AND ITS EFFECT ON MASCULINITY AND FEMININITY IN PREDICATIVE COMPOUND