EP1717668A1 - Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporating the same - Google Patents
Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporating the same Download PDFInfo
- Publication number
- EP1717668A1 EP1717668A1 EP05252711A EP05252711A EP1717668A1 EP 1717668 A1 EP1717668 A1 EP 1717668A1 EP 05252711 A EP05252711 A EP 05252711A EP 05252711 A EP05252711 A EP 05252711A EP 1717668 A1 EP1717668 A1 EP 1717668A1
- Authority
- EP
- European Patent Office
- Prior art keywords
- language
- text
- new
- electronic device
- handheld electronic
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Withdrawn
Links
Images
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F3/00—Input arrangements for transferring data to be processed into a form capable of being handled by the computer; Output arrangements for transferring data from processing unit to output unit, e.g. interface arrangements
- G06F3/01—Input arrangements or combined input and output arrangements for interaction between user and computer
- G06F3/02—Input arrangements using manually operated switches, e.g. using keyboards or dials
- G06F3/023—Arrangements for converting discrete items of information into a coded form, e.g. arrangements for interpreting keyboard generated codes as alphanumeric codes, operand codes or instruction codes
- G06F3/0233—Character input methods
- G06F3/0237—Character input methods using prediction or retrieval techniques
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F40/00—Handling natural language data
- G06F40/20—Natural language analysis
- G06F40/237—Lexical tools
- G06F40/242—Dictionaries
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F40/00—Handling natural language data
- G06F40/20—Natural language analysis
- G06F40/263—Language identification
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F40/00—Handling natural language data
- G06F40/20—Natural language analysis
- G06F40/274—Converting codes to words; Guess-ahead of partial word inputs
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04M—TELEPHONIC COMMUNICATION
- H04M1/00—Substation equipment, e.g. for use by subscribers
- H04M1/72—Mobile telephones; Cordless telephones, i.e. devices for establishing wireless links to base stations without route selection
- H04M1/724—User interfaces specially adapted for cordless or mobile telephones
- H04M1/72403—User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality
- H04M1/7243—User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality with interactive means for internal management of messages
- H04M1/72436—User interfaces specially adapted for cordless or mobile telephones with means for local support of applications that increase the functionality with interactive means for internal management of messages for text messaging, e.g. short messaging services [SMS] or e-mails
Definitions
- aspects of the invention relate to generating text in a handheld electronic device and to expediting the process, such as for example, where the handheld electronic device receives text from sources external to the device.
- Generating text in a handheld electronic device examples of which include, for instance, personal data assistants (PDA's), handheld computers, two-way pagers, cellular telephones, text messaging devices, and the like, has become a complex process. This is due at least partially to the trend to make these handheld electronic devices smaller and lighter in weight. A limitation in making them smaller has been the physical size of keyboard if the keys are to be actuated directly by human fingers. Generally, there have been two approaches to solving this problem. One is to adapt the ten digit keypad indigenous to mobile phones for text input. This requires each key to support input of multiple characters. The second approach seeks to shrink the traditional full keyboard, such as the "qwerty" keyboard by doubling up characters to reduce the number of keys.
- An object of aspects of the invention is to facilitate generating text in a handheld electronic device.
- an object is to assist the generation of text by processes that utilize lists of words, ideograms and the like by gathering new language objects from sources of text external to the handheld electronic device.
- the generation of text in a handheld electronic device that utilizes lists of language objects, such as for example, words, abbreviations, text shortcuts, and in some languages ideograms and the like to facilitate text generation, adapts to the user's experience by adding new language objects gleaned from text received from sources external to the handheld electronic device.
- An exemplary external source of text is e-mail messages. Additional non-limiting examples include SMS (Short Message Service), MMS (Multi-Media Service) and instant messages.
- aspects of the invention are directed to a method of entering text into a handheld electronic device.
- the handheld electronic device has at least one application for receiving text from sources external to the handheld electronic device and a text input process that accesses at least one list of stored language objects to facilitate generation of text.
- the general nature of the method can be stated as including processing received text received from an external source comprising scanning the received text for any new language objects not in any list of stored language objects, and identifying any of the new language objects that fail to meet a number of specified characteristics that are at least partially determinative of a language.
- aspects of the invention also embrace a handheld electronic device having a plurality of applications including at least one that receives text from a source external to the handheld electronic device.
- the device also includes a user interface through which a user inputs linguistic elements and a text generator that has a first language object list and a new language object list and a text input processor.
- This text input processor comprises processing means selecting new language objects not in the first or new list and identifying any of the new language objects that fail to meet a number of specified characteristics that are at least partially determinative of a language, and means using selected language objects stored in the first list and the new list to generate the desired text from the linguistic elements input through the user interface.
- This handheld electronic device also includes an output means presenting the desired text to the user.
- FIG. 1 illustrates a wireless handheld electronic device 1, which is but one type of handheld electronic device to which aspects of the invention can be applied.
- the exemplary handheld electronic device 1 includes an input device 3 in the form of a keyboard 5 and a thumbwheel 7 that are used to control the functions of the handheld electronic device 1 and to generate text and other inputs.
- the keyboard 5 constitutes a reduced "qwerty" keyboard in which most of the keys 9 are used to input two letters of the alphabet. Thus, initially the input generated by depressing one of these keys is ambiguous in that it is undetermined as to which letter was intended.
- Various schemes have been devised for disambiguating the inputs generated by these keys 9 assigned multiple letters for input. The particular scheme used is not relevant to aspects of the invention as long as one or more linguistic lists are used in the process.
- the input provided through the keyboard 5 and thumbwheel 7 are displayed on a display 11 as is well known.
- the input device 3 provides keystroke inputs to an execution system 13 that may be an operating system, a Java virtual machine, a run time environment or the like.
- the handheld electronic device 1 implements a plurality of applications 17. These applications can include an address book 19, e-mail 21, a calendar 23, a memo 25, and additional applications such as, for example, spell check and a phone application. Generally these applications 17 require text input that is implemented by a text input process 27, which forms part of an input system 15.
- Various types of text input processes 27 can be used that employ lists 29 to facilitate the generation of text.
- the text input process 27 utilizes software to progressively narrow the possible combination of letters that could be intended by a specified sequence of keystrokes.
- Such "disambiguation" software is known.
- Such systems employ a plurality of lists of linguistic objects.
- linguistic objects it is meant in the example words and in some languages ideograms.
- the keystrokes input linguistic elements which in the case of words, are characters or letters in the alphabet, and in the case of ideograms, strokes that make up the ideogram.
- the list of language objects can also include abbreviations, and text shortcuts, which are becoming common with the growing use of various kinds of text messaging.
- Text shortcuts embraces the cryptic and rather clever short representations of common messages, such as, for example, "CUL8R” for “see you later”, “PXT” for “please explain that", "SS” for “so sorry”, and the like.
- Lists that can be used by the exemplary disambiguation text input process 27 can include a generic list 31 and a new list 33. Additional lists 35 can include learned words and special word lists such as technical terms for biotechnology.
- Known disambiguation programs can assign frequencies of use to the language objects, such as words, in the lists it uses to determine the language object intended by the user. Frequencies of use can be initially assigned based on statistics of common usage and can then be modified through actual usage. It is known for disambiguation programs to incorporate "learned" language objects such as words that were not in the initial lists, but were inserted by the user to drive the output to the intended new word. It is known to assign such learned words an initial frequency of use that is near the high end of the range of frequencies of use. This initial frequency of use is then modified through actual use as with the initially inserted words.
- aspects of the present invention are related to increasing the language objects available for use by the text input process 27.
- One source for such additional language objects is the e-mail application. Not only is it likely that new language objects contained in incoming e-mails would be used by the user to generate a reply or other e-mail responses, such new language objects could also be language objects that the user might want to use in generating other text inputs.
- Figures 3 and 4 illustrate a flow chart of a routine 38 for harvesting new language objects from received e-mails.
- the incoming e-mails 39 are placed in a queue 41 for processing as permitted by the processing burden on the handheld electronic device 1.
- Processing begins with scanning the e-mail to parse the message into words (language objects) at 43.
- the parsed message is then filtered at 45 to remove unwanted components, such as numbers, dates, and the like.
- the language objects are then compared with the language objects in the current lists at 47. If it is determined at 49 that none of the language objects in the received text are missing from the current lists, such as if all of the language objects in the incoming e-mail message are already in one of the lists as determined at 47, then the routine 38 returns to the queue at 41.
- the text input process then initiates scanning of the next incoming e-mail in the queue as processing time becomes available.
- processing continues to 51 where it is determined whether any of the new language objects can be considered to be in the current language being employed by the user on the handheld electronic device 1 to input text.
- An example of the processing at 51 is described in greater detail in Figure 4 and below. If it is determined at 51 that no new language objects are in the current language, all of the new language objects are ignored, and the routine returns to the queue at 41. If, however, it is determined at 51 that a new language object is in the current language, each such new language object in the current language is assigned a frequency of use at 53.
- This assigned frequency of use will typically be in the high range of the frequencies of use, for the example, at about the top one third.
- These new words are placed in the new list 33. However, such a list will have a certain finite capacity, such that over time the new list can become full, as determined at 55. If such is the case, room must be made for this latest entry. Thus, at 57, room is made in the new list by removing one of the earlier entries.
- the new words are assigned a selected high initial frequency of use, and that frequency of use diminishes through operation of the disambiguation routine of the text input process
- the word with the lowest frequency of use can be removed from the new list to make room for the latest new word.
- the stored new language object having a time stamp that is oldest can be removed. Accordingly, this latest new word is added to the new list at 59 and the routine returns to the queue at 41.
- An exemplary language analysis procedure such as is performed at 51, is depicted in detail in Figure 4. It is first determined whether the ratio of new language objects in at least a segment of the text to the total number of language objects in the segment exceeds a predetermined threshold. For instance, if an analysis were performed on the text on a line-by-line basis, the routine 38 would determine at 61 whether the quantity of new language objects in any line of text is, for example, ten percent (10%) or more of the quantity of language objects in the line of text. Any appropriate threshold may be employed. Also, segments of the text other than lines may be analyzed, or the entire text message can be analyzed as a whole. The size of the segment may be determined based upon the quantity of text in the message and/or upon other factors. If it is determined at 61 that the threshold has not been met, the new language objects in the text are accepted as being in the current language, and processing continues onward to 53, as is indicated at the numeral 69 in Figure 4.
- processing continues at 63 where the linguistic elements in all of the new language objects in the text are compared with a set of predetermined linguistic elements.
- a determination of the ratio of new language objects to language objects and the set of predetermined linguistic elements are non-limiting examples of specified characteristics that may be at least partially indicative of or particular to one or more predetermined languages.
- an exemplary set of predetermined linguistic elements indicative of the English language might include, for instance, the twenty-six Latin letters, both upper and lower case, symbols such as an ampersand, asterisk, exclamation point, question mark, and pound sign, and certain predetermined diacritics. If a new language object has a linguistic element other than the linguistic elements in the set of predetermined linguistic elements particular to the current language, the new language object is considered to be in a language other than the current language. If the English language is the current language used on the handheld electronic device 1, such as if the language objects stored in the lists 29 are generally in the English language, the routine 38 can identify and ignore non-English words.
- any new language objects are identified at 63 as having a linguistic element not in the set of predetermined linguistic elements, such new language objects are ignored, as at 65.
- the routine 38 determines at 67 whether any non-ignored new language objects exist in the text. If yes, the routine 38 then ascertains at 68 whether a ratio of the ignored new language objects in the text to the new language objects in the text exceeds another threshold, for example fifty percent (50%). Any appropriate threshold may be applied. For instance, if the routine 38 determines at 68 that fifty percent or more of the new language objects were ignored at 65, processing returns to the queue at 41, as is indicated at the numeral 71 in Figure 4. This can provide an additional safeguard against adding undesirable language objects to the new list 33.
- routine 38 determines at 68 that fewer than fifty percent of the new language objects were ignored at 65, processing continues at 53, as is indicated in Figure 4 at the numeral 69, where the non-ignored new language objects can be added to the new list 33.
- processing returns to the queue at 41 as is indicated in Figure 4 at the numeral 71. It is understood that other language analysis methodologies may be employed.
- the above process not only searches for new words in a received e-mail but also for new abbreviations and new text shortcuts, or for ideograms if the language uses ideograms.
- other text received from sources outside the handheld electronic device can also be scanned for new words. This can include gleaning new language objects from instant messages, SMS (short message service), MMS (multimedia service), and the like.
Landscapes
- Engineering & Computer Science (AREA)
- Theoretical Computer Science (AREA)
- General Engineering & Computer Science (AREA)
- Physics & Mathematics (AREA)
- General Physics & Mathematics (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Health & Medical Sciences (AREA)
- Artificial Intelligence (AREA)
- Computational Linguistics (AREA)
- General Health & Medical Sciences (AREA)
- Human Computer Interaction (AREA)
- Business, Economics & Management (AREA)
- General Business, Economics & Management (AREA)
- Computer Networks & Wireless Communication (AREA)
- Signal Processing (AREA)
- Telephone Function (AREA)
- Information Transfer Between Computers (AREA)
Abstract
Description
- Aspects of the invention relate to generating text in a handheld electronic device and to expediting the process, such as for example, where the handheld electronic device receives text from sources external to the device.
- Generating text in a handheld electronic device examples of which include, for instance, personal data assistants (PDA's), handheld computers, two-way pagers, cellular telephones, text messaging devices, and the like, has become a complex process. This is due at least partially to the trend to make these handheld electronic devices smaller and lighter in weight. A limitation in making them smaller has been the physical size of keyboard if the keys are to be actuated directly by human fingers. Generally, there have been two approaches to solving this problem. One is to adapt the ten digit keypad indigenous to mobile phones for text input. This requires each key to support input of multiple characters. The second approach seeks to shrink the traditional full keyboard, such as the "qwerty" keyboard by doubling up characters to reduce the number of keys. In both cases, the input generated by actuation of a key representing multiple characters is ambiguous. Various schemes have been devised to interpret inputs from these multi-character keys. Some schemes require actuation of the key a specific number of times to identify the desired character. Others use software to progressively narrow the possible combinations of letters that can be intended by a specified sequence of key strokes. This latter approach uses multiple lists that can contain, for instance, generic words, application specific words, learned words and the like.
- An object of aspects of the invention is to facilitate generating text in a handheld electronic device. In another sense, an object is to assist the generation of text by processes that utilize lists of words, ideograms and the like by gathering new language objects from sources of text external to the handheld electronic device.
- The generation of text in a handheld electronic device that utilizes lists of language objects, such as for example, words, abbreviations, text shortcuts, and in some languages ideograms and the like to facilitate text generation, adapts to the user's experience by adding new language objects gleaned from text received from sources external to the handheld electronic device. An exemplary external source of text is e-mail messages. Additional non-limiting examples include SMS (Short Message Service), MMS (Multi-Media Service) and instant messages.
- More particularly, aspects of the invention are directed to a method of entering text into a handheld electronic device. The handheld electronic device has at least one application for receiving text from sources external to the handheld electronic device and a text input process that accesses at least one list of stored language objects to facilitate generation of text. The general nature of the method can be stated as including processing received text received from an external source comprising scanning the received text for any new language objects not in any list of stored language objects, and identifying any of the new language objects that fail to meet a number of specified characteristics that are at least partially determinative of a language.
- Aspects of the invention also embrace a handheld electronic device having a plurality of applications including at least one that receives text from a source external to the handheld electronic device. The device also includes a user interface through which a user inputs linguistic elements and a text generator that has a first language object list and a new language object list and a text input processor. This text input processor comprises processing means selecting new language objects not in the first or new list and identifying any of the new language objects that fail to meet a number of specified characteristics that are at least partially determinative of a language, and means using selected language objects stored in the first list and the new list to generate the desired text from the linguistic elements input through the user interface. This handheld electronic device also includes an output means presenting the desired text to the user.
-
- Figure 1 is a front view of an exemplary handheld electronic device incorporating aspects of the invention.
- Figure 2 is a functional diagram in block form illustrating aspects of the invention.
- Figure 3 is a flow chart illustrating operation of aspects of the invention.
- Figure 4 is a flow chart illustrating operation of aspects of the invention.
- Figure 1 illustrates a wireless handheld
electronic device 1, which is but one type of handheld electronic device to which aspects of the invention can be applied. The exemplary handheldelectronic device 1 includes aninput device 3 in the form of akeyboard 5 and athumbwheel 7 that are used to control the functions of the handheldelectronic device 1 and to generate text and other inputs. Thekeyboard 5 constitutes a reduced "qwerty" keyboard in which most of thekeys 9 are used to input two letters of the alphabet. Thus, initially the input generated by depressing one of these keys is ambiguous in that it is undetermined as to which letter was intended. Various schemes have been devised for disambiguating the inputs generated by thesekeys 9 assigned multiple letters for input. The particular scheme used is not relevant to aspects of the invention as long as one or more linguistic lists are used in the process. The input provided through thekeyboard 5 andthumbwheel 7 are displayed on adisplay 11 as is well known. - Turning to Figure 2, the
input device 3 provides keystroke inputs to anexecution system 13 that may be an operating system, a Java virtual machine, a run time environment or the like. The handheldelectronic device 1 implements a plurality ofapplications 17. These applications can include anaddress book 19,e-mail 21, acalendar 23, amemo 25, and additional applications such as, for example, spell check and a phone application. Generally theseapplications 17 require text input that is implemented by atext input process 27, which forms part of aninput system 15. - Various types of
text input processes 27 can be used thatemploy lists 29 to facilitate the generation of text. For example, in the exemplary handheld electronic device where the reduced "qwerty" keyboard produces ambiguous inputs, thetext input process 27 utilizes software to progressively narrow the possible combination of letters that could be intended by a specified sequence of keystrokes. Such "disambiguation" software is known. Typically, such systems employ a plurality of lists of linguistic objects. By linguistic objects it is meant in the example words and in some languages ideograms. The keystrokes input linguistic elements, which in the case of words, are characters or letters in the alphabet, and in the case of ideograms, strokes that make up the ideogram. The list of language objects can also include abbreviations, and text shortcuts, which are becoming common with the growing use of various kinds of text messaging. Text shortcuts embraces the cryptic and rather clever short representations of common messages, such as, for example, "CUL8R" for "see you later", "PXT" for "please explain that", "SS" for "so sorry", and the like. Lists that can be used by the exemplary disambiguationtext input process 27 can include ageneric list 31 and anew list 33.Additional lists 35 can include learned words and special word lists such as technical terms for biotechnology. Other types oftext input processes 27, such as for example, prediction programs that anticipate a word intended by a user as it is typed in and thereby complete it, could also use word lists. Such a prediction program might be used with a full keyboard. - Known disambiguation programs can assign frequencies of use to the language objects, such as words, in the lists it uses to determine the language object intended by the user. Frequencies of use can be initially assigned based on statistics of common usage and can then be modified through actual usage. It is known for disambiguation programs to incorporate "learned" language objects such as words that were not in the initial lists, but were inserted by the user to drive the output to the intended new word. It is known to assign such learned words an initial frequency of use that is near the high end of the range of frequencies of use. This initial frequency of use is then modified through actual use as with the initially inserted words.
- Aspects of the present invention are related to increasing the language objects available for use by the
text input process 27. One source for such additional language objects is the e-mail application. Not only is it likely that new language objects contained in incoming e-mails would be used by the user to generate a reply or other e-mail responses, such new language objects could also be language objects that the user might want to use in generating other text inputs. - Figures 3 and 4 illustrate a flow chart of a routine 38 for harvesting new language objects from received e-mails. The
incoming e-mails 39 are placed in aqueue 41 for processing as permitted by the processing burden on the handheldelectronic device 1. Processing begins with scanning the e-mail to parse the message into words (language objects) at 43. The parsed message is then filtered at 45 to remove unwanted components, such as numbers, dates, and the like. The language objects are then compared with the language objects in the current lists at 47. If it is determined at 49 that none of the language objects in the received text are missing from the current lists, such as if all of the language objects in the incoming e-mail message are already in one of the lists as determined at 47, then the routine 38 returns to the queue at 41. The text input process then initiates scanning of the next incoming e-mail in the queue as processing time becomes available. - However, if any of the language objects examined at 47 are determined at 49 to be missing from the current lists, meaning that they are new language objects, processing continues to 51 where it is determined whether any of the new language objects can be considered to be in the current language being employed by the user on the handheld
electronic device 1 to input text. An example of the processing at 51 is described in greater detail in Figure 4 and below. If it is determined at 51 that no new language objects are in the current language, all of the new language objects are ignored, and the routine returns to the queue at 41. If, however, it is determined at 51 that a new language object is in the current language, each such new language object in the current language is assigned a frequency of use at 53. This assigned frequency of use will typically be in the high range of the frequencies of use, for the example, at about the top one third. These new words are placed in thenew list 33. However, such a list will have a certain finite capacity, such that over time the new list can become full, as determined at 55. If such is the case, room must be made for this latest entry. Thus, at 57, room is made in the new list by removing one of the earlier entries. In the exemplary embodiment, where the new words are assigned a selected high initial frequency of use, and that frequency of use diminishes through operation of the disambiguation routine of the text input process, the word with the lowest frequency of use can be removed from the new list to make room for the latest new word. Alternatively, the stored new language object having a time stamp that is oldest can be removed. Accordingly, this latest new word is added to the new list at 59 and the routine returns to the queue at 41. - An exemplary language analysis procedure, such as is performed at 51, is depicted in detail in Figure 4. It is first determined whether the ratio of new language objects in at least a segment of the text to the total number of language objects in the segment exceeds a predetermined threshold. For instance, if an analysis were performed on the text on a line-by-line basis, the routine 38 would determine at 61 whether the quantity of new language objects in any line of text is, for example, ten percent (10%) or more of the quantity of language objects in the line of text. Any appropriate threshold may be employed. Also, segments of the text other than lines may be analyzed, or the entire text message can be analyzed as a whole. The size of the segment may be determined based upon the quantity of text in the message and/or upon other factors. If it is determined at 61 that the threshold has not been met, the new language objects in the text are accepted as being in the current language, and processing continues onward to 53, as is indicated at the numeral 69 in Figure 4.
- On the other hand, continuing the example, if it is determined at 61 that in any line or other segment of text the threshold is exceeded, processing continues at 63 where the linguistic elements in all of the new language objects in the text are compared with a set of predetermined linguistic elements. A determination of the ratio of new language objects to language objects and the set of predetermined linguistic elements are non-limiting examples of specified characteristics that may be at least partially indicative of or particular to one or more predetermined languages.
- If, for example, the current language is English, an exemplary set of predetermined linguistic elements indicative of the English language might include, for instance, the twenty-six Latin letters, both upper and lower case, symbols such as an ampersand, asterisk, exclamation point, question mark, and pound sign, and certain predetermined diacritics. If a new language object has a linguistic element other than the linguistic elements in the set of predetermined linguistic elements particular to the current language, the new language object is considered to be in a language other than the current language. If the English language is the current language used on the handheld
electronic device 1, such as if the language objects stored in thelists 29 are generally in the English language, the routine 38 can identify and ignore non-English words. - If any new language objects are identified at 63 as having a linguistic element not in the set of predetermined linguistic elements, such new language objects are ignored, as at 65. The routine 38 then determines at 67 whether any non-ignored new language objects exist in the text. If yes, the routine 38 then ascertains at 68 whether a ratio of the ignored new language objects in the text to the new language objects in the text exceeds another threshold, for example fifty percent (50%). Any appropriate threshold may be applied. For instance, if the routine 38 determines at 68 that fifty percent or more of the new language objects were ignored at 65, processing returns to the queue at 41, as is indicated at the numeral 71 in Figure 4. This can provide an additional safeguard against adding undesirable language objects to the
new list 33. On the other hand, if the routine 38 determines at 68 that fewer than fifty percent of the new language objects were ignored at 65, processing continues at 53, as is indicated in Figure 4 at the numeral 69, where the non-ignored new language objects can be added to thenew list 33. - If it is determined at 67 that no non-ignored new language objects exist in the text, processing returns to the queue at 41 as is indicated in Figure 4 at the numeral 71. It is understood that other language analysis methodologies may be employed.
- The above process not only searches for new words in a received e-mail but also for new abbreviations and new text shortcuts, or for ideograms if the language uses ideograms. In addition to scanning e-mails for new words, other text received from sources outside the handheld electronic device can also be scanned for new words. This can include gleaning new language objects from instant messages, SMS (short message service), MMS (multimedia service), and the like.
- While specific embodiments of the invention have been described in detail, it will be appreciated by those skilled in the art that various modifications and alternatives to those details could be developed in light of the overall teachings of the disclosure. Accordingly, the particular arrangements disclosed are meant to be illustrative only and not limiting as to the scope of the invention which is to be given the full breadth of the claims appended and any and all equivalents thereof.
Claims (20)
- A method of entering text into a handheld electronic device having at least one application for receiving text from sources external to the handheld electronic device and a text input process that accesses at least one list of stored language objects to facilitate generation of text, the method comprising:processing received text received from an external source comprising scanning the received text for any new language objects not in any list of stored language objects; andidentifying any of the new language objects that fail to meet a number of specified characteristics that are at least partially determinative of at least a first predetermined language.
- The method of Claim 1, wherein said identifying comprises determining that a ratio of the quantity of new language objects in at least a segment of the received text with the quantity of language objects in the at least a segment of the received text exceeds a predetermined threshold.
- The method of Claim 2, further comprising, responsive to said determining, identifying at least a first new language object on the basis that the at least a first new language object is in a language other than the at least a first predetermined language.
- The method of Claim 3 wherein each new language object comprises a number of linguistic elements, wherein the handheld electronic device has a list of predetermined linguistic elements that correspond with the at least a first predetermined language, and wherein said identifying at least a first new language object comprises determining that the at least a first new language object comprises at least a first linguistic element different than the predetermined linguistic elements in the list.
- The method of Claim 1, wherein said identifying comprises identifying at least a first new language object on the basis that the at least a first new language object is in a language other than at least a first predetermined language.
- The method of Claim 5 wherein each new language object comprises a number of linguistic elements, wherein the handheld electronic device has a list of predetermined linguistic elements that correspond with the at least a first predetermined language, and wherein said identifying at least a first new language object comprises determining that the at least a first new language object comprises at least a first linguistic element different than the predetermined linguistic elements in the list.
- The method of Claim 2, further comprising determining that in at least a first segment, the quantity of new language objects compared with the quantity of language objects exceeds a predetermined threshold.
- The method of Claim 7, further comprising identifying at least a first new language object on the basis that the at least a first new language object is in a language other than at least a first predetermined language
- The method of Claim 8, further comprising determining that a ratio of identified new language objects to new language objects exceeds a predetermined threshold and, responsive thereto, ignoring all of the new language objects.
- The method of Claim 1, further comprising adding any new language object that has not been identified to the at least one list of stored language objects for use by the text input process in generating text.
- A handheld electronic device comprising:a plurality of applications that utilize text and at least one of which receives text from a source external to the handheld electronic device;a user interface through which a user inputs linguistic elements for generating text;a text generator comprising:a first list storing language objects;a new list storing new language objects; anda text input processor comprising means selecting from among language objects in received text from the source external to the handheld electronic device new language objects not in the first list or the new list and identifying any of the new language objects that fail to meet a number of specified characteristics that are at least partially determinative of a language, and means using selected language objects stored in the first list and the new list to generate the desired text from the linguistic elements input through the user interface; andoutput means presenting the desired text to the user.
- The handheld electronic device of Claim 11 wherein the text generator is adapted to determine that a ratio of the quantity of new language objects in at least a segment of the received text with the quantity of language objects in the at least a segment of the received text exceeds a predetermined threshold.
- The handheld electronic device of Claim 12 wherein the text generator is adapted to identify at least a first new language object on the basis that the at least a first new language object is in a language other than at least a first predetermined language.
- The handheld electronic device of Claim 13 wherein each new language object comprises a number of linguistic elements, wherein the handheld electronic device has a list of predetermined linguistic elements that correspond with the at least a first predetermined language, and wherein the text generator is adapted to determine that the at least a first new language object comprises at least a first linguistic element different than the predetermined linguistic elements in the list.
- The handheld electronic device of Claim 11, wherein the text generator is adapted to identify at least a first new language object on the basis that the at least a first new language object is in a language other than at least a first predetermined language.
- The handheld electronic device of Claim 15 wherein each new language object comprises a number of linguistic elements, wherein the handheld electronic device has a list of predetermined linguistic elements that correspond with the at least a first predetermined language, and wherein the text generator is adapted to determine that the at least a first new language object comprises at least a first linguistic element different than the predetermined linguistic elements in the list.
- The handheld electronic device of Claim 12 wherein the text generator is adapted to determine that in at least a first segment, the quantity of new language objects compared with the quantity of language objects exceeds a predetermined threshold.
- The handheld electronic device of Claim 17 wherein the text generator is adapted to identify at least a first new language object on the basis that the at least a first new language object is in a language other than at least a first predetermined language
- The handheld electronic device of Claim 18 wherein the text generator is adapted to determine that a ratio of identified new language objects to new language objects exceeds a predetermined threshold and, responsive thereto, the text generator is adapted to ignore all of the new language objects.
- The handheld electronic device of Claim 11 wherein the text generator is adapted to add any new language object that has not been identified to the at least one list of stored language objects for use by the text input process in generating text.
Priority Applications (5)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| EP05252711A EP1717668A1 (en) | 2005-04-29 | 2005-04-29 | Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporating the same |
| CA2606328A CA2606328C (en) | 2005-04-29 | 2006-04-25 | Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporating the same |
| PCT/CA2006/000660 WO2006116845A1 (en) | 2005-04-29 | 2006-04-25 | Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporating the same |
| GB0723227A GB2443337B (en) | 2005-04-29 | 2006-04-25 | Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporatingthe same |
| DE112006001079T DE112006001079T5 (en) | 2005-04-29 | 2006-04-25 | A method of generating text that satisfies specified characteristics in a portable electronic device and a portable electronic device including the same |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| EP05252711A EP1717668A1 (en) | 2005-04-29 | 2005-04-29 | Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporating the same |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| EP1717668A1 true EP1717668A1 (en) | 2006-11-02 |
Family
ID=34941124
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP05252711A Withdrawn EP1717668A1 (en) | 2005-04-29 | 2005-04-29 | Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporating the same |
Country Status (5)
| Country | Link |
|---|---|
| EP (1) | EP1717668A1 (en) |
| CA (1) | CA2606328C (en) |
| DE (1) | DE112006001079T5 (en) |
| GB (1) | GB2443337B (en) |
| WO (1) | WO2006116845A1 (en) |
Cited By (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP1965594A1 (en) * | 2007-03-01 | 2008-09-03 | Huawei Technologies Co., Ltd. | Method for processing short message and communication terminal |
| WO2008120043A1 (en) * | 2007-03-29 | 2008-10-09 | Nokia Corporation | Method, apparatus, system, user interface and computer program product for use with managing content |
| EP2081119A1 (en) | 2008-01-17 | 2009-07-22 | Research In Motion Limited | Obtaining new language objects for a temporary dictionary used by a disambiguation routine |
| EP2286350A4 (en) * | 2008-06-06 | 2012-08-29 | Zi Corp Canada Inc | Systems and methods for an automated personalized dictionary generator for portable devices |
Citations (8)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO1997005541A1 (en) * | 1995-07-26 | 1997-02-13 | King Martin T | Reduced keyboard disambiguating system |
| GB2318659A (en) * | 1996-09-30 | 1998-04-29 | Ibm | Natural language determination using letter pairs in commonly occuring words |
| EP1255184A2 (en) * | 2001-05-04 | 2002-11-06 | Nokia Corporation | A communication terminal having a predictive text editor application |
| US6553103B1 (en) * | 2000-07-20 | 2003-04-22 | International Business Machines Corporation | Communication macro composer |
| EP1320023A2 (en) * | 2001-11-27 | 2003-06-18 | Nokia Corporation | A communication terminal having a text editor application |
| US20030233235A1 (en) * | 2002-06-17 | 2003-12-18 | International Business Machines Corporation | System, method, program product, and networking use for recognizing words and their parts of speech in one or more natural languages |
| GB2396940A (en) * | 2002-12-31 | 2004-07-07 | Nokia Corp | A predictive text editor utilising words from received text messages |
| EP1480421A1 (en) * | 2003-05-20 | 2004-11-24 | Sony Ericsson Mobile Communications AB | Automatic setting of a keypad input mode in response to an incoming text message |
Family Cites Families (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US5953541A (en) * | 1997-01-24 | 1999-09-14 | Tegic Communications, Inc. | Disambiguating system for disambiguating ambiguous input sequences by displaying objects associated with the generated input sequences in the order of decreasing frequency of use |
| DE60304246T2 (en) * | 2003-05-20 | 2006-11-02 | Sony Ericsson Mobile Communications Ab | Setting the mode selection depending on language information |
-
2005
- 2005-04-29 EP EP05252711A patent/EP1717668A1/en not_active Withdrawn
-
2006
- 2006-04-25 WO PCT/CA2006/000660 patent/WO2006116845A1/en not_active Ceased
- 2006-04-25 DE DE112006001079T patent/DE112006001079T5/en not_active Withdrawn
- 2006-04-25 CA CA2606328A patent/CA2606328C/en not_active Expired - Lifetime
- 2006-04-25 GB GB0723227A patent/GB2443337B/en not_active Expired - Lifetime
Patent Citations (8)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO1997005541A1 (en) * | 1995-07-26 | 1997-02-13 | King Martin T | Reduced keyboard disambiguating system |
| GB2318659A (en) * | 1996-09-30 | 1998-04-29 | Ibm | Natural language determination using letter pairs in commonly occuring words |
| US6553103B1 (en) * | 2000-07-20 | 2003-04-22 | International Business Machines Corporation | Communication macro composer |
| EP1255184A2 (en) * | 2001-05-04 | 2002-11-06 | Nokia Corporation | A communication terminal having a predictive text editor application |
| EP1320023A2 (en) * | 2001-11-27 | 2003-06-18 | Nokia Corporation | A communication terminal having a text editor application |
| US20030233235A1 (en) * | 2002-06-17 | 2003-12-18 | International Business Machines Corporation | System, method, program product, and networking use for recognizing words and their parts of speech in one or more natural languages |
| GB2396940A (en) * | 2002-12-31 | 2004-07-07 | Nokia Corp | A predictive text editor utilising words from received text messages |
| EP1480421A1 (en) * | 2003-05-20 | 2004-11-24 | Sony Ericsson Mobile Communications AB | Automatic setting of a keypad input mode in response to an incoming text message |
Non-Patent Citations (2)
| Title |
|---|
| KENNETH R. BEESLEY: "Language Identifier: A Computer Program for Automatic Natural-Language Identification of On-Line Text", AUTOMATED LANGUAGE PROCESSING SYSTEMS, pages 1 - 21, XP002343517, Retrieved from the Internet <URL:http://www.xrce.xerox.com/competencies/content-analysis/tools/publis/langid.pdf> [retrieved on 20050905] * |
| PROCEEDINGS OF THE 29TH ANNUAL CONFERENCE OF THE AMERICAN TRANSLATORS ASSOCIATION, 12 October 1988 (1988-10-12) - 16 October 1988 (1988-10-16), pages 47 - 54 * |
Cited By (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP1965594A1 (en) * | 2007-03-01 | 2008-09-03 | Huawei Technologies Co., Ltd. | Method for processing short message and communication terminal |
| US8019366B2 (en) | 2007-03-01 | 2011-09-13 | Huawei Technologies Co., Ltd | Method for processing short message and communication terminal |
| WO2008120043A1 (en) * | 2007-03-29 | 2008-10-09 | Nokia Corporation | Method, apparatus, system, user interface and computer program product for use with managing content |
| EP2081119A1 (en) | 2008-01-17 | 2009-07-22 | Research In Motion Limited | Obtaining new language objects for a temporary dictionary used by a disambiguation routine |
| EP2286350A4 (en) * | 2008-06-06 | 2012-08-29 | Zi Corp Canada Inc | Systems and methods for an automated personalized dictionary generator for portable devices |
Also Published As
| Publication number | Publication date |
|---|---|
| CA2606328C (en) | 2012-04-24 |
| GB2443337B (en) | 2010-10-13 |
| DE112006001079T5 (en) | 2008-03-20 |
| GB0723227D0 (en) | 2008-01-09 |
| GB2443337A (en) | 2008-04-30 |
| WO2006116845A1 (en) | 2006-11-09 |
| CA2606328A1 (en) | 2006-11-09 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US8554544B2 (en) | Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporating the same | |
| US9851983B2 (en) | Method for generating text in a handheld electronic device and a handheld electronic device incorporating the same | |
| US20130191112A1 (en) | Method for identifying language of text in a handheld electronic device and a handheld electronic device incorporating the same | |
| US7580829B2 (en) | Apparatus and method for reordering of multiple language databases for text disambiguation | |
| KR100893447B1 (en) | Contextual prediction of user words and user actions | |
| US7562007B2 (en) | Method and apparatus for recognizing language input mode and method and apparatus for automatically switching language input modes using the same | |
| EP2133772B1 (en) | Device and method incorporating an improved text input mechanism | |
| US20080182599A1 (en) | Method and apparatus for user input | |
| EP2109046A1 (en) | Predictive text input system and method involving two concurrent ranking means | |
| KR20100046043A (en) | Disambiguation of keypad text entry | |
| CA2605777C (en) | Method for generating text in a handheld electronic device and a handheld electronic device incorporating the same | |
| CA2606328C (en) | Method for generating text that meets specified characteristics in a handheld electronic device and a handheld electronic device incorporating the same | |
| US20130013296A1 (en) | Handheld electronic device with reduced keyboard and associated method of providing improved disambiguation with reduced degradation of device performance | |
| CN101169686A (en) | Stroke input method | |
| CA2661559C (en) | Method for identifying language of text in a handheld electronic device and a handheld electronic device incorporating the same | |
| CA2605785C (en) | Handheld electronic device with reduced keyboard and associated method of providing improved disambiguation with reduced degradation of device performance | |
| EP1632872A1 (en) | System and method for managing databases in a handheld electronic device | |
| US20060047628A1 (en) | System and method for managing databases in a handheld electronic device | |
| Sunny | Text entry for mobile devices | |
| HK1080192B (en) | Method and apparatus for inputting ideographic characters into devices | |
| HK1087503A (en) | System and method for managing databases in a handheld electronic device |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| PUAI | Public reference made under article 153(3) epc to a published international application that has entered the european phase |
Free format text: ORIGINAL CODE: 0009012 |
|
| 17P | Request for examination filed |
Effective date: 20050520 |
|
| AK | Designated contracting states |
Kind code of ref document: A1 Designated state(s): AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HU IE IS IT LI LT LU MC NL PL PT RO SE SI SK TR |
|
| AX | Request for extension of the european patent |
Extension state: AL BA HR LV MK YU |
|
| 17Q | First examination report despatched |
Effective date: 20070124 |
|
| AKX | Designation fees paid |
Designated state(s): AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HU IE IS IT LI LT LU MC NL PL PT RO SE SI SK TR |
|
| AXX | Extension fees paid |
Extension state: YU Payment date: 20070419 Extension state: MK Payment date: 20070419 Extension state: LV Payment date: 20070419 Extension state: HR Payment date: 20070419 Extension state: BA Payment date: 20070419 Extension state: AL Payment date: 20070419 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE APPLICATION HAS BEEN WITHDRAWN |
|
| 18W | Application withdrawn |
Effective date: 20100126 |