BTemplates.com

Powered by Blogger.

Pageviews past week

Quantum mechanics

Auto News

artificial intelligence

About Me

Recommend us on Google!

Information Technology

Popular Posts

Showing posts with label Web browser. Show all posts
Showing posts with label Web browser. Show all posts

Thursday, June 30, 2011

Can Google Get Web Users Talking?


Voice-driven search is a futuristic idea, and may take some getting used to.
Credit: Google

The notion of asking a computer for information out loud is familiar to most of us only from science fiction. Google is trying to change that by adding speech recognition to its search engine, and releasing technology that would allow any browser, website, or app to use the feature.

But are you ready to give up your keyboards and talk to Google instead?

Over the last two weeks, speech input for Google has gradually been rolled out to every person using Google's Chrome browser. A microphone icon appears at the right end of the iconic search box. If you have a microphone built-in or attached to your computer, clicking that icon creates a direct audio connection to Google's servers, which will convert your spoken words into text.

It has been possible to speak Google search queries using a smart phone for almost three years; since last year, Android handsets have been able to take voice input in any situation where a keyboard would normally be used. "That was transformational, because people stopped worrying about when they could and couldn't speak to the phone," says Vincent Vanhoucke, who leads the voice search engineering team at Google. Over the last 12 months, the number of spoken inputs, search or otherwise, via Android devices has climbed six times, and every day, tens of thousands of hours of audio speech are fed into Google's servers. "On Android, a large fraction of the use is people dictating e-mail and SMS," says Vanhoucke.

Vanhoucke's team now wants using voice on the Web to be as easy as it is on Android. "It's a big bet," he says. "Voice search for desktop is the flagship for this, [but] we want to take speech everywhere."

Voice recognition is more technically challenging on a desktop or laptop computer, says Vanhoucke, because it requires noise suppression algorithms that are not needed for mobile speech recognition. These algorithms filter out sounds such as those of a computer's fan or air conditioners. "The quality of the audio is paramount for phone manufacturers, and you hold it close to your mouth," says Vanhoucke. "On a PC, the microphone is an afterthought, and you are further away. You don't get the best quality."



Google asked thousands of people to read phrases aloud to their computers to gather data on the conditions its speech recognition technology would have to handle. As people use the service for real, it is trained further, says Vanhoucke, which should increase its popularity. Data from users of mobile voice search shows that people are much more likely to use the feature again when it is accurate for them the first time.

A bigger challenge to getting users to embrace voice recognition on the desktop could be the existing tools for entering information, says Keith Vertanen, a lecturer at Princeton University who researches voice-recognition technology. "On the desktop, you're up against a very fast and efficient means of input in the keyboard," he says. "On a phone, you don't have that available, and you are often in hands- or eyes-free situations where voice input really helps."

Vertanen says people are less tolerant of glitches when using speech recognition on a desktop computer because of the close proximity of a tried-and-true way of entering text. He says users might find voice recognition more compelling on on other Internet-connected devices in the home. "Nonconventional devices like a DVR, television, or game console don't usually have good text input," he points out. Google TV devices can already take voice input spoken into a connected Android phone.

Vanhoucke acknowledges that speech recognition fulfills a more immediate need on phones, but argues that users are ready for it on conventional computers, too. "People will use it in ways that surprise us," he says. "At this point, it's still an experiment." Situations when people may have their hands full is one example, says Vanhoucke (although it should be noted that desktop voice search today still involves using the mouse to activate the feature).

Google isn't performing this experiment alone. The company is pushing the Web standards body W3C to introduce a standard set of HTML markup that allows any website or app to call on voice recognition via the Web browser, and has already enabled a version of this markup in the Chrome browser. For now, Google is the only major company with a browser able to use the prototype feature, but Mozilla, Microsoft, and AT&T are all working with the W3C effort.

"It's a collaborative effort that other browser makers are part of," says Vanhoucke. "Any designer can add it to their Web page. It's something anyone can use." Extensions for the Chrome browser that make use of voice input (like this one) have already appeared, and can be used to enter text on any website.

However, those extensions reveal that although Google's desktop speech recognition is accurate for search queries, it's not much good for tasks like composing e-mail.

Enabling the system to learn the personal quirks of each person's pronunciation, a feature already enabled on Android phones, could address that. Vertanen points out that the personalization learned through mobile search could easily be ported over to the desktop for people logged into their Google account. It could also make it possible for the technology to spring up elsewhere. "The advantage of Google's networked approach is that a [speech] model in the cloud can adapt to your voice in all these different places and follow you around, whether that's in your living room or in your car."


A Browser that Speaks Your Language The latest version of Google's Chrome shows the potential of HTML5.

Sunday, April 10, 2011

A Browser that Speaks Your Language The latest version of Google's Chrome shows the potential of HTML5.


Early adopters can now get a sneak peek at the future of the Web by downloading the latest prerelease, or "beta," version of Chrome, Google's Web browser. One of the most interesting new features is an ability to translate speech to text—entirely via the Web.
Credit: Technology Review

The feature is the result of work Google has been doing with the World Wide Web Consortium's HTML Speech Incubator Group, the mission of which is "to determine the feasibility of integrating speech technology in HTML5," the Web's new, emerging standard language.

A Web page employing the new HTML5 feature could have an icon that, when clicked, initiates a recording through the computer's microphone, via the browser. Speech is captured and sent to Google's servers for transcription, and the resulting text is sent back to the website.

To experiment with the voice-to-text feature, download the latest beta version of Chrome here. Then go to this webpage, click on the microphone, and start talking. You'll probably find the results mixed, and sometimes hilarious. Using the finest elocution I could muster, I read the opening passage of Richard Yates's Revolutionary Road: "The final dying sounds of their dress rehearsal left the Laurel Players with nothing to do but stand there, silent and helpless." I got error messages several times in a row ("speech not recognized" or "connection to speech servers failed"). Once, I received this transcription: "9 sounds good restaurants on the world there's nothing to do with fam vans island."

The new feature derives in large part from experiments Google conducted through its Android operating system for mobile devices. For more than a year, says Vincent Vanhoucke, a member of Google's voice recognition team, Android app developers have been able to integrate voice recognition into their apps using technology provided by Google. This has provided Google with useful voice data with which to train its voice-recognition algorithms. Today, some 20 percent of searches on Android phones are conducted using voice recognition, says Vanhoucke: people use voice recognition to write texts, send emails, or conduct searches. "It has really opened up interesting new avenues," says Vanhoucke.

However, unlike desktop voice-to-text software, which first accustoms itself to a user's voice, Chrome is trying to churn out text from voice without prior training.

"I suppose if they keep track of [the] IP address, they could adapt" to a given user's voice, says Jim Glass, a speech recognition expert at MIT. Glass notes that the mobile phone provides an acoustic environment very different from that of a laptop or desktop computer; for one thing, a phone's microphone is reliably placed right at the user's mouth, unlike computer microphone setups in homes or offices. "This is the beta version of Chrome," says Glass. "They'll be collecting data, and we can be sure they will be refining their models--that's the nature of the speech-recognition game."

Even if it's rough around the edges, sometimes the technology impresses. I tried once again and got back "the final warning sounds of the dress rehearsal at laurel players with nothing to do with stand there." Not so bad. And the Chrome app nailed it to a letter when all I said was "the quick brown fox jumps over the lazy dog."

Third-party programmers have also begun creating Web pages capable of using the new feature of Chrome. Already available for trial is a browser plugin called Speechify that lets you search Google, Hulu, YouTube, Amazon, and other sites using voice with Chrome.

Other inventive uses could soon follow. "Games could be taking keyboard, mouse, touch, accelerometer, and speech input together," says Karl Westin, an expert on HTML5 who works for Nerd Communications, based in Berlin, Germany. "Having an aeroplane game where you could actually scream 'up, UP, UUUPPP!' could be fantastic."

But the technology is more than just a toy—it also points the way to a much more capable Web. HTML4, the last major version of the HTML language, emerged in 1997. Since then, plugins like Silverlight and Flash have added media-processing capabilities to the Web. But HTML5 enables media playback and offline storage via the browser.

"The insight we had was that more and more people were spending all their time in the browser," says Google's Brian Rakowski, group product manager for Chrome. E-mail and instant messaging increasingly take place in browsers rather than in separate e-mail or AIM applications. "We'd like it to be case that you never have to install a native application again," says Rakowski. "The Web should be able to do all of it."

Saturday, February 6, 2010

Liquid Phone with Qualcomm Snapdragon Processor


Acer has brought a new liquid phone in India. It’s really liquid due to its unique feature which differentiates it from the other phones. This phone has the world’s first Qualcomm Snapdragon processor and is based on the first Android 1.6 high definition smart phone. It delivers real time communication as well as content which are location aware. This smart device brings forth a unique combination of high quality performance as well as bold style.

This high defining smart phone combines the cutting edge technologies; software innovation as well as ultra-fluid user interfaces so that users can get a completely new experience with this phone. Unique set of features developed by Acer and its partners would include a new user interface as well as improved power management systems. This enhanced power management system would help the user to achieve longer battery autonomy.

Saturday, February 28, 2009

Apple announces Safari 4


Apple announces the launch of the world’s fastest and most innovative browser – Safari 4


Apple announced the public beta of Safari 4, the world’s fastest and most innovative web browser for Mac and Windows PCs. The Nitro engine in Safari 4 runs JavaScript 4.2 times faster than Safari 3. Innovative new features that include top sites, for a stunning visual preview of frequently visited pages; full history search, to search through titles, etc, make browsing more intuitive and enjoyable. Mistake

“Apple created Safari to bring innovation, speed and open standards back into web browsers, and today it takes another big step forward,” said Philip Schiller, Apple’s senior vice president of Worldwide Product Marketing. “Safari 4 is the fastest and most efficient browser for Mac and Windows, with great integration of HTML 5 and CSS 3 web standards that enables the next generation of interactive web applications.”

Safari 4 is built on the world’s most advanced browser technologies including the new Nitro JavaScript engine that executes JavaScript up to 30 times faster than IE 7 and more than three times faster than Firefox 3. Safari quickly loads HTML web pages three times faster than IE 7 and almost three times faster than Firefox 3.

Safari 4 includes HTML 5 support for offline technologies so web-based applications can store information locally without an Internet connection, and is the first browser to support advanced CSS Effects that enable highly polished web graphics using reflections, gradients and precision masks. Safari 4 is the first browser to pass the Web Standards Project’s Acid3 test, which examines how well a browser adheres to CSS, JavaScript, XML and SVG web standards that are specifically designed for dynamic web applications.

Safari for Mac, Windows, iPhone and iPod touch are all built on Apple’s WebKit, the world’s fastest and most advanced browser engine. Apple developed WebKit as an open source project to create the world’s best browser engine and to advance the adoption of modern web standards. Recently, WebKit led the introduction of HTML 5 and CSS 3 web standards and is known for its fast, modern code-base. The industry’s newest browsers are based on WebKit including Google Chrome, the Google Android browser, the Nokia Series 60 browser and Palm webOS.

New features in Safari 4 include:

• Top Sites, a display of frequently visited pages in a stunning wall of previews so users can jump to their favorite sites with a single click;

• Full History Search, where users search through titles, web addresses and the complete text of recently viewed pages to easily return to sites they’ve seen before;

• Cover Flow, to make searching web history or bookmarks as fun and easy as paging through album art in iTunes®;

• Tabs on Top, for better tabbed browsing with easy drag-and-drop tab management tools and an intuitive button for opening new ones;

• Smart Address Field, that automatically completes web addresses by displaying an easy-to-read list of suggestions from Top Sites, bookmarks and browsing history;

• Smart Search Field, where users fine-tune searches with recommendations from Google Suggest or a list of recent searches;

• Full Page Zoom, for a closer look at any website without degrading the quality of the site’s layout and text;

• built-in web developer tools to debug, tweak and optimize a website for peak performance and compatibility; and

• a new Windows-native look in Safari for Windows, that uses standard Windows font rendering and native title bar, borders and toolbars so Safari fits the look and feel of other Windows XP and Windows Vista applications.

Pricing and Availability

Safari 4 is a public beta for both Mac OS X and Windows and is available immediately as a free download at www.apple.com/safari.

Safari 4 for Mac OS X requires Mac OS X Leopard version 10.5.6 and Security Update 2009-001 or Mac OS X Tiger version 10.4.11, a minimum 256MB of memory, and is designed to run on any Intel-based Mac or a Mac with a PowerPC G5, G4 or G3 processor and built-in FireWire. Safari 4 for Windows requires Windows XP SP2 or Windows Vista, a minimum 256MB of memory and a system with at least a 500 MHz Intel Pentium processor. Full system requirements and more information on Safari 4 can be found at www.apple.com/safari

Reblog this post [with Zemanta]