By using this site, you agree to the Privacy Policy and Terms of Use.
Accept

GadgetBond

  • Latest
  • How-to
  • Tech
    • AI
    • Amazon
    • Apple
    • CES
    • Computing
    • Creators
    • Google
    • Meta
    • Microsoft
    • Mobile
    • Samsung
    • Security
    • Xbox
  • Transportation
    • Audi
    • BMW
    • Cadillac
    • E-Bike
    • Ferrari
    • Ford
    • Honda Prelude
    • Lamborghini
    • McLaren W1
    • Mercedes
    • Porsche
    • Rivian
    • Tesla
  • Culture
    • Apple TV
    • Disney
    • Gaming
    • Hulu
    • Marvel
    • HBO Max
    • Netflix
    • Paramount
    • SHOWTIME
    • Star Wars
    • Streaming
Add GadgetBond as a preferred source to see more of our stories on Google.
Font ResizerAa
GadgetBondGadgetBond
  • Latest
  • Tech
  • AI
  • Deals
  • How-to
  • Apps
  • Mobile
  • Gaming
  • Streaming
  • Transportation
Search
  • Latest
  • Deals
  • How-to
  • Tech
    • Amazon
    • Apple
    • CES
    • Computing
    • Creators
    • Google
    • Meta
    • Microsoft
    • Mobile
    • Samsung
    • Security
    • Xbox
  • AI
    • Anthropic
    • ChatGPT
    • ChatGPT Atlas
    • Gemini AI (formerly Bard)
    • Google DeepMind
    • Grok AI
    • Meta AI
    • Microsoft Copilot
    • OpenAI
    • Perplexity
    • xAI
  • Transportation
    • Audi
    • BMW
    • Cadillac
    • E-Bike
    • Ferrari
    • Ford
    • Honda Prelude
    • Lamborghini
    • McLaren W1
    • Mercedes
    • Porsche
    • Rivian
    • Tesla
  • Culture
    • Apple TV
    • Disney
    • Gaming
    • Hulu
    • Marvel
    • HBO Max
    • Netflix
    • Paramount
    • SHOWTIME
    • Star Wars
    • Streaming
Follow US
AIOpenAITech

OpenAI’s GPT-4 now understands both text and image inputs

By
Shubham Sawarkar
Shubham Sawarkar's avatar
ByShubham Sawarkar
Editor-in-Chief
I’m a tech enthusiast who loves exploring gadgets, trends, and innovations. With certifications in CISCO Routing & Switching and Windows Server Administration, I bring a sharp...
Follow:
- Editor-in-Chief
Mar 15, 2023, 8:47 AM EDT
Share
We may get a commission from retail offers. Learn more
OpenAI's GPT-4 now understands both text and image inputs
(Photo by D koi on Unsplash)
SHARE

OpenAI has announced the launch of GPT-4, the latest iteration of its generative pre-trained transformer system. Unlike its predecessor GPT-3.5, which can only read and respond with text, GPT-4 can generate text on input images. This development comes hot on the heels of Google’s Workspace AI announcement and ahead of Microsoft’s Future of Work event. OpenAI has reportedly spent the past six months refining the system’s performance based on user feedback generated from the recent ChatGPT conversational bot hype. The company claims that GPT-4 exhibits human-level performance on various professional and academic benchmarks.

OpenAI has partnered with Microsoft to develop GPT’s capabilities and has achieved record performance in “factuality, steerability, and refusing to go outside of guardrails” compared to its predecessor. The new system has also outperformed other state-of-the-art large language models (LLMs) in a variety of benchmark tests. GPT-4 will be made available for both ChatGPT and the API, but access will be restricted to ChatGPT Plus subscribers and API waitlist users, respectively. There will also be a usage cap in place for playing with the new model.

The added multi-modal input feature of GPT-4 will generate text outputs based on a wide variety of mixed text and image inputs. This means that users can scan marketing and sales reports, textbooks, shop manuals, and even screenshots, and ChatGPT will summarize the various details into small words that are easy to understand. The recently upgraded system can be customized by the API developer, allowing developers and soon ChatGPT users to prescribe their AI’s style and task by describing those directions in the ‘system’ message.

GPT-4 has been tested by 50 experts in a wide array of professional fields, and the model’s tendency to “hallucinate” facts has been reduced by around 40 percent compared to its predecessor. The new model is also 82 percent less likely to respond to requests for disallowed content. However, OpenAI still strongly recommends that great care should be taken when using language model outputs, particularly in high-stakes contexts, and that the exact protocol should match the needs of a specific use case.

OpenAI’s GPT-4 represents a significant advance in AI technology, with its ability to generate text on input images and improved performance in various benchmarks. As AI continues to develop, it will be interesting to see how GPT-4 and other systems like it can be used in practical applications.


Discover more from GadgetBond

Subscribe to get the latest posts sent to your email.

Leave a Comment

Leave a ReplyCancel reply

Most Popular

Perplexity Computer is now inside Microsoft Teams

Apple gives up on Vision Pro after M5 refresh fails

Google Docs now lets you set custom instructions for Gemini

Apple’s rumored 32-inch iMac Ultra sounds absolutely wild

Google Workspace now has a central hub to control all AI and agent access

Also Read
Perplexity illustration. Abstract illustration of a transparent glass cube refracting beams of light into rainbow-like streaks across a dark, textured surface, symbolizing clarity, synthesis, and the convergence of multiple perspectives.

Perplexity Agent API now ships with Finance Search for structured financial insight

Apple showing off Siri’s updated logo at WWDC 2024.

Apple faces $250 million payout after overselling AI Siri on iPhone 16

The OpenAI logo displayed in white against a deep blue gradient background. The logo consists of a stylized hexagonal geometric shape resembling an interlocking pattern or aperture on the left, paired with the text "OpenAI" in a clean, modern font on the right. The background features subtle lighting effects with darker edges and a brighter blue glow in the upper right corner, creating a professional and technological atmosphere.

OpenAI’s rumored ChatGPT phone targets 2027 launch window

Minimal promotional graphic featuring the text “GPT-5.5 Instant” centered inside a rounded white rectangle, set against a soft abstract background with blurred pastel gradients in pink, purple, orange, and blue tones.

GPT-5.5 Instant replaces GPT-5.3 as OpenAI’s everyday ChatGPT model

Promotional interface mockup for Perplexity Computer focused on professional finance workflows, showing an “NVDA Post Earnings Impact Memo” with financial tables, charts, and analysis sections alongside a task panel requesting an AI-generated NVIDIA earnings summary with market insights and semiconductor industry implications.

Perplexity launches Computer for Professional Finance

Abstract 3D illustration of a flowing metallic ribbon with reflective gold and silver surfaces, curved in a wave-like shape against a dark background with bright light reflections and glossy highlights.

Perplexity health search gets a major upgrade with Premium Sources

Illustration of Google Chrome enhanced autofill showing three side-by-side form examples for loyalty card numbers, vehicle license plates, and travel confirmation numbers. Each input field displays a dropdown suggestion card with saved information and management options against a blue background.

Google Chrome’s enhanced autofill completely changes how you fill out tedious online forms

Close-up of the Google Drive webpage showing the Drive logo, the heading “Drive,” and text about storing, accessing, and sharing files, with a “Get started” button visible.

Google Drive API now supports large-scale CSE file migrations

Company Info
  • Homepage
  • Support my work
  • Latest stories
  • Company updates
  • GDB Recommends
  • Daily newsletters
  • About us
  • Contact us
  • Write for us
  • Editorial guidelines
Legal
  • Privacy Policy
  • Cookies Policy
  • Terms & Conditions
  • DMCA
  • Disclaimer
  • Accessibility Policy
  • Security Policy
  • Do Not Sell or Share My Personal Information
Socials
Follow US

Disclosure: We love the products we feature and hope you’ll love them too. If you purchase through a link on our site, we may receive compensation at no additional cost to you. Read our ethics statement. Please note that pricing and availability are subject to change.

Copyright © 2026 GadgetBond. All Rights Reserved. Use of this site constitutes acceptance of our Terms of Use and Privacy Policy | Do Not Sell/Share My Personal Information.

Advertisement
Amazon Summer Beauty Event 2026