User interface selectable real time information delivery system and method
An information delivery system including a client device and server interconnected by a network passes data files in accordance with a well known protocol. The server creates an audio version of a displayed page associated with a first data file. The audio version includes information from the first data file and other items, so that the information is presented in the form of conversation-like natural speech. The server merges the audio version and the data file to produce a second data file and delivers the second data file to the client device. Advantageously, the client device includes a speech synthesizer engine and a display so that the data file can be viewed using the first data file and/or heard by a user using the audio version, both included in the second data file.
This application is a divisional of application Ser. No. 10/274,685, filed Oct. 21, 2002, which is a continuation of and claims the benefit of U.S. application Ser. No. 10/081,159 filed on Feb. 21, 2002, which claims priority from U.S. Provisional Application No. 60,270,358 filed on Feb. 21, 2001. The above mentioned applications are incorporated herein by reference.
FIELD OF THE INVENTIONThe present invention generally relates to an information delivery system, and more particularly to a system that generates a user-friendly version of a first data file, wherein the first data file is suited for one type of user interface and the user-friendly version is suited for a different type of user interface. Both versions are simultaneously delivered to a user, so that the user can retrieve the same information from two different types of user interface.
BACKGROUND OF THE INVENTIONThe Internet provides a robust facility for providing information on diverse topics. For many topics, such as account information and stock quotes, the information consists primarily of tables or lists of numbers and symbols, and usually in a format that is suited only for a graphic display in a user device, such as a monitor attached to a computer. Generally, a service provider does not provide a voice translation of the displayed information. Thus, the user has no option to listen to such information even if the user device is equipped with a speaker.
A possible solution is to use a text-to-speech converter. However, unlike news story for example, this type of information is not in the format of straight text, i.e., not in the form of conversation-like natural speech or acontextual. As a result, the converted audio may be incomprehensible. Thus, there is a need to develop a system and method for enabling a user to listen to such information in the form of conversation-like audio.
SUMMARY OF THE INVENTIONA system according to the principles of the invention enables users to retrieve information from different types of user interfaces. The information is originally saved in a format suitable for a particular type of user interface, such as video displays. The information is then converted to a different format suitable for a different type of user interface, such as an audio speaker. The converted format includes the information provided in the original format but also includes other elements, so that the information retrieved by the different type of user interface is tailored to natural human communication. For example, if the different type of user interface is an audio speaker, prefatory and other transitional phases may be added to communicate the information in a manner most closely resembling natural language speech.
A system according to the principles of the invention includes a client device connected to an information server via a network wherein the client device and server are adapted to pass data files (such as hypertext files) in accordance with a well known protocol (such as HyperText Transfer Protocol—“HTTP”). The server is further adapted to create data files suitable for a first type of user interface from data files suitable for a second but different type of user interface as requested by the client device. Such data files can be created in real-time and may contain either real-time information and/or historical information.
The system allows a user accessing the server via the client device to request and view the requested data files on the display of the client device. To provide this capability, the data file received by the client device is read using a well known markup language reader, such as a web or wap browser. Advantageously, the present invention further includes a speech synthesis engine installed on the client device adapted to convert information from the data file into an audio format.
In another embodiment of the present invention, the server is adapted to deliver data files containing information along with settings for controlling the operation of the speech synthesizer engine in the client device.
In another embodiment of the present invention, the server includes a storage device for storing the control settings of the speech synthesizer engine.
In yet another embodiment of the present invention, the server delivers data files containing information in both an audible and a visual format. The hypertext files further include a user interface for selecting access to the information in audible format.
In yet another embodiment, the server and client device may be configured to allow the client device to deliver unsolicited information in a data file. The user may pre-select whether the delivered information is provided in a visual and/or an audible format.
BRIEF DESCRIPTION OF THE DRAWINGSA more complete understanding of the present invention may be obtained from consideration of the following description in conjunction with the drawings in which:
With reference to the
The server is further adapted to create HTML data files in real time as requested by the client device. Accordingly, a user, accessing the server via the client device, can request and view these real time HTML data files on display 11 using a HTML reader installed in the client device. The HTML reader passes the information in the HTML files to the graphic designated by reference numeral 12 in
Advantageously, the present invention further includes a speech synthesis engine 14 installed on the client device. This engine is adapted to convert information from a HTML file into an audio format and comprises two layers. The first layer interfaces with an HTML compatible software program, such as a web browser, and retrieves any HTML file having an audio component from the browser and generates an audio output therefrom. The audio output is preferably in the form of text. The second layer is any speech application program interface (“SAPI”) compatible program. One example of a SAPI compatible program is SAPI Version 5.0 distributed by Microsoft Corporation of Redmond, Wash., U.S.A. Reference numeral 14 indicates the audio output to speaker 13 of client device 10. If the audio output is in the form of text, a text-to-speech converter (not shown) included in the speech engine can be used to pronounce the text.
In another embodiment of the present invention, the server is adapted to deliver hypertext files containing information and control settings for operating the speech engine contained on the client device.
The server can include storage 32 for storing the control settings of the speech synthesizer.
In yet another embodiment of the present invention, the server delivers HTML files containing information in both an audible and a visual format. The hypertext files further include a user interface for selecting access to the information in an audible format.
The server and client device can be configured to provide for the delivery of unsolicited information in a HTML file. The unsolicited information as used herein is the information delivered to a user other than in response to an interactive request. Rather, the user may subscribe to a service available in the server and the server then delivers the information provided by that service when certain events have occurred. For example, the user may subscribe to a service that periodically supplies updated stock quotes for certain stocks selected by the user in an interval specified by the user. The user may pre-select whether the information is to be delivered to the user in a visual or an audible format. More particularly, the server and the client device are adapted to deliver hypertext files to the user wherein the information in such files is tailored to the user.
With reference to
The present invention may be used on a server adapted to transmit information to a remote user in real time. While the information can include any kind of information; in the preferred embodiment of the invention, the server includes stock quote, transaction and current client account information, including real time reporting of user-selected stock market indicators. For example,
The HTML rendering engine is responsive to a client device having a text-to-speech engine 14. In such case, the information delivered to the client device includes commands and information tailored for an audible format. The audio message will vary according to user, time of day and other real time information. An implementation may have the server application merging standard templates with customized user, data source and time of day information. When the HTML file is received, the information formatted for visual display is displayed by the client device using a conventional HTTP compatible browser and the information formatted for audio output is delivered to speech engine 14, which uses a text-to-speech converter.
The use of the text-to-speech engine allows for the reporting of information to the user without requiring this user to focus on the display. Thus, the user can do other tasks away from the client device or operate their account in the background while doing other tasks.
It should be noted that information is delivered in a different way when delivered by audible format rather than by visual format. For example, additional prefatory phrases or other transitional phrases not required for a visual format are required for the audible format to communicate the information in a manner most closely resembling natural language speech. Natural language speech, for purposes of this application, is not limited to any particular natural language, e.g., English, German, French, etc., but refers to any natural language.
As an example,
The content of some of the variable items depends on user preferences set by the user through, for example, the customization page shown in
The text shown in
Referring to
The present invention is particularly well suited for use when monitoring information for particular content, such as waiting for a particular transaction to occur. Delivered information may be broadcast to the user upon delivery. Therefore, if the user is not at the client device but within hearing distance of the client device's audio output device, information can still be effectively communicated to the user.
The methods described above can also be implemented in a computer readable medium without deviating from the principles of the invention.
Numerous modifications and alternative embodiments of the invention will be apparent to those skilled in the art in view of the foregoing description. Accordingly, this description is to be construed as illustrative only and is for the purpose of teaching those skilled in the art the best mode of carrying out the invention. Details of the structure may be varied substantially without departing from the spirit of the invention and the exclusive use of all modifications which come within the scope of the appended claim is reserved.
Claims
1. A method for delivering information from a server across a network in an audible format of a displayed page, the method comprising the steps of:
- creating a template for the audible format of the displayed page, the template including fixed and variable items;
- filling the variable items in the template using data at least from the displayed page based on a user preferences set; and
- delivering data in the template with the filled variable items to a client device of a user using a network communication protocol.
2. The method of claim 1 further comprising the step of receiving a request from the user before the filling step.
3. The method of claim 1 further comprising the step of receiving a request from the user before the delivering step.
4. The method of claim 1, wherein the variable items include user preferences saved in a database.
5. The method of claim 4, wherein the filling step includes the step of retrieving the user preferences from the database for filling the variable items.
6. A method for displaying a page and playing an audio version of the page in a client device, the method comprising the steps of:
- receiving the audio version, wherein the audio version is created from a template having fixed and variable items, and the variable items are filled with data at least from the page;
- converting the audio version into text; and
- playing the text using a text-to-speech converter.
7. The method of claim 6, wherein the playing step is performed by a speech engine in the client device and the audio version includes control settings for the speech engine.
8. The method of claim 7, wherein one of the settings specifies a playing speed in the unit of words per second.
9. A server for delivering information across a network in an audible format of a displayed page, the server comprising:
- a first data base for storing a template for the audible format of the displayed page, the template including fixed and variable items; and
- a rendering engine for filling the variable items in the template using data at least from the displayed page based on a user preferences set and delivering data in the template to a client device of a user using a network communication protocol.
10. The server of claim 9 further comprising a second database for storing user preferences.
11. The server of claim 10 wherein the rendering machine retrieves one of the user preferences for filling one of the variable items in the template.
12. A client device for displaying a page and playing an audio version of the page in a client device, the client device comprising:
- a remote client for receiving the audio version, wherein the audio version is created from a template having fixed and variable items, and the variable items are filled with data at least from the page based on a user preferences set; and
- a speech engine playing the audio version using a text-to-speech converter.
13. The device of claim 12 wherein the speech engine is configured by control settings included in the audio version.
14. The device of claim 13, wherein one of the settings specifies a playing speed in the unit of words per second.
15. A method for providing natural speech rendering of acontextual data, the method comprising:
- in a structure comprising natural speech phrases and variable items, filling the variable items with corresponding acontextual data from a database having time associated updates, wherein the natural speech phrases provide context to the acontextual data;
- rendering a combination of the natural speech phrases and acontextual data.
Type: Application
Filed: Jan 20, 2006
Publication Date: Jun 1, 2006
Inventors: Hans Gilde (Brooklyn, NY), Steele Arbeeny (Guttenberg, NJ)
Application Number: 11/336,182
International Classification: G10L 21/00 (20060101);