Researchers develop AI to make the internet more accessible

By News EditorPublished On: January 10, 2024Last Updated: January 10, 2024

In an effort to make the internet more accessible for people with disabilities, researchers at The Ohio State University have begun developing an artificial intelligence (AI) agent that could complete complex tasks on any website using simple language commands.

In the three decades since it was first released into the public domain, the internet has become an incredibly intricate, dynamic system.

However, because internet function is now so integral to society’s well-being, its complexity also makes it considerably harder to navigate.

Today there are billions of websites available to help access information or communicate with others, and many online tasks can take more than a dozen steps to complete.

That’s why Yu Su, co-author of the study and an assistant professor of computer science and engineering at Ohio State, said their work, which uses information taken from live sites to create web agents — online AI helpers — is a step towards making the digital world a less confusing place.

Su said: “For some people, especially those with disabilities, it’s not easy for them to browse the internet.

“We rely more and more on the computing world in our daily life and work, but there are increasingly a lot of barriers to that access, which, to some degree, widens the disparity.”

The study work was presented in December at the Thirty-seventh Conference on Neural Information Processing Systems (NeurIPS), a flagship conference for AI and machine learning research.

By taking advantage of the power of large language models, the agent works similarly to how humans behave when browsing the web, Su said.

The Ohio State researchers showed that their model was able to understand the layout and functionality of different websites using only its ability to process and predict language.

They started the process by creating Mind2Web, the first dataset for generalist web agents.

Though previous efforts to build web agents focused on toy simulated websites, Mind2Web fully embraces the complex and dynamic nature of real-world websites and emphasises an agent’s ability of generalising to entirely new websites it has never seen before.

Much of their success is due to their agent’s ability to handle the internet’s ever-evolving learning curve, Su said.

The researchers lifted over 2,000 open-ended tasks from 137 different real-world websites, which they then used to train the agent.

Some of the tasks included booking one-way and round-trip international flights, following celebrity accounts on Twitter, browsing comedy films from 1992 to 2017 streaming on Netflix, and even scheduling car knowledge tests at the US Department of Motor Vehicles (DMV).

Many of the tasks were very complex, such as booking one of the international flights used in the model would take 14 actions.

Such effortless versatility allows for diverse coverage on a number of websites, opening up a new landscape for future models to explore and learn in an autonomous fashion, said Su.

Su said: “It’s only become possible to do something like this because of the recent development of large language models like ChatGPT.”

Since the chatbot became public in November 2022, millions of users have used it to automatically generate content, ranging from poetry and jokes to cooking advice and medical diagnoses.

But because one website could contain thousands of raw HTML elements, it would be too costly to feed so much information to a single large language model.

To address the gap, the study also introduces a framework called MindAct, a two-pronged agent that uses both small and large language models to carry out these tasks.

The researchers found that by using this strategy, MindAct significantly outperforms other common modeling strategies and is able to understand various concepts at a decent level.

With more fine-tuning, the study said, the model could likely be used in tandem with both open-and closed-source large language models such as Flan-T5 or GPT-4.

However, their research does highlight an increasingly relevant ethical problem in creating flexible artificial intelligence, said Su.

While it could certainly serve as a helpful agent to humans surfing the web, the model could also be used to enhance systems like ChatGPT and turn the entire internet into an unprecedentedly powerful tool, Su said.

He said: “On the one hand, we have great potential to improve our efficiency and to allow us to focus on the most creative part of our work.

“But on the other hand, there’s tremendous potential for harm.”

For example, autonomous agents able to translate online steps into the real world could influence society by taking potentially dangerous actions, such as misusing financial information or spreading misinformation.

Su said: “We should be extremely cautious about these factors and make a concerted effort to try to mitigate them.”

But as AI research continues to evolve, he said that it’s likely society will experience major growth in the commercial use and performance of generalist web agents in the years to come, especially as the technology has already gained so much popularity in the public eye.

Su added: “Throughout my career, my goal has always been trying to bridge the gap between human users and the computing world.

“That said, the real value of this tool is that it will really save people time and make the impossible possible.”

How automated triage changes the game for clinicians dealing with trauma referrals

AI could help reduce alcohol-related risks in surgery patients

Cookie	Duration	Description
__cfduid	1 month	The cookie is used by cdn services like CloudFare to identify individual clients behind a shared IP address and apply security settings on a per-client basis. It does not correspond to any user ID in the web application and does not store any personally identifiable information.
__hssrc	session	This cookie is set by Hubspot. According to their documentation, whenever HubSpot changes the session cookie, this cookie is also set to determine if the visitor has restarted their browser. If this cookie does not exist when HubSpot manages cookies, it is considered a new session.
cookielawinfo-checkbox-advertisement	1 year	The cookie is set by GDPR cookie consent to record the user consent for the cookies in the category "Advertisement".
cookielawinfo-checkbox-analytics	1 year	This cookies is set by GDPR Cookie Consent WordPress Plugin. The cookie is used to remember the user consent for the cookies under the category "Analytics".
cookielawinfo-checkbox-necessary	1 year	This cookie is set by GDPR Cookie Consent plugin. The cookies is used to store the user consent for the cookies in the category "Necessary".
cookielawinfo-checkbox-performance	1 year	This cookie is set by GDPR Cookie Consent plugin. The cookie is used to store the user consent for the cookies in the category "Performance".

Cookie	Duration	Description
__hssc	30 minutes	This cookie is set by HubSpot. The purpose of the cookie is to keep track of sessions. This is used to determine if HubSpot should increment the session number and timestamps in the __hstc cookie. It contains the domain, viewCount (increments each pageView in a session), and session start timestamp.
tve_leads_unique	1 month	This cookie is set by the provider Thrive Themes. This cookie is used to know which optin form the visitor has filled out when subscribing a newsletter.

Cookie	Duration	Description
__hstc	1 year 24 days	This cookie is set by Hubspot and is used for tracking visitors. It contains the domain, utk, initial timestamp (first visit), last timestamp (last visit), current timestamp (this visit), and session number (increments for each subsequent session).
_ga	2 years	This cookie is installed by Google Analytics. The cookie is used to calculate visitor, session, campaign data and keep track of site usage for the site's analytics report. The cookies store information anonymously and assign a randomly generated number to identify unique visitors.
_gid	1 day	This cookie is installed by Google Analytics. The cookie is used to store information of how visitors use a website and helps in creating an analytics report of how the wbsite is doing. The data collected including the number visitors, the source where they have come from, and the pages viisted in an anonymous form.
hubspotutk	1 year 24 days	This cookie is used by HubSpot to keep track of the visitors to the website. This cookie is passed to Hubspot on form submission and used when deduplicating contacts.

Cookie	Duration	Description
cookielawinfo-checkbox-functional	1 year	The cookie is set by GDPR cookie consent to record the user consent for the cookies in the category "Functional".
cookielawinfo-checkbox-others	1 year	No description
lfuuid	9 years 11 months	Third party (Lead Forensics) cookie which enables us to track visitor behaviour on our site. Tracking is performed anonymously until a user identifies themselves by submitting a form.
tl_554_555_1	1 month	No description
tl_554_605_2	1 month	No description
tlf_1	5 days	No description