Tag: Vaizdų generavimas

  • „OpenAI“, „Meta“ ir „xAI“ paleidžia naujus DI modelius: kas keisis vartotojams jau šią savaitę

    „OpenAI“, „Meta“ ir „xAI“ paleidžia naujus DI modelius: kas keisis vartotojams jau šią savaitę

    By&nbspRoselyne Min
    Published on

    Share

    Comments

    Add Euronews on Google

    Share
    Close Button












    It’s a week of major AI releases from some of the world’s biggest tech companies. Here’s what to expect.

    Several major technology companies are releasing new AI models or updates this week, as the AI race promises to be hotter than the summer heatwave.


    ADVERTISEMENT


    ADVERTISEMENT

    OpenAI is expected to make GPT-5.6 publicly available on Thursday, widening access to its latest and most capable AI model after an earlier limited rollout.

    The company first unveiled GPT-5.6 in late June, but access was restricted to a small group of vetted partners while the US government reviewed potential national security risks.

    Similarly to the release, withdrawal and subsequent re-release of Anthropic’s Fable model, US officials had raised concerns that increasingly powerful AI systems could be misused, including for cyber or military purposes.

    OpenAI said in a post on X late Tuesday that it now plans to release GPT-5.6 Sol, along with its Terra and Luna models. Sol is the company’s most advanced model, while Terra is a lower-cost mid-tier option and Luna is its most cost-efficient version.

    The precarious rollout schedule by leading AI companies also highlights how the race is no longer merely about capability, but also about who controls deployment, whose data powers the tools and where — and how — they are used.

    The delayed rollout reflects growing government scrutiny of frontier AI systems, as policymakers seek more oversight of models that could be used in sensitive areas such as cybersecurity, defence and intelligence.

    In June, the Trump administration signed an executive order establishing a voluntary framework under which AI developers could offer “covered frontier models” to the US government for up to 30 days before releasing them to trusted partners, and later, to the wider public.

    “We don’t believe this kind of government access process should become the long-term default. It keeps the best tools from users, developers, enterprises, cyber defenders, and global partners who need them,” OpenAI wrote in an announcement in June.

    “We are taking this short-term step because we believe it is the strongest path to broader availability in the coming weeks, while we work with the Administration to develop the cyber Executive Order framework and a repeatable process for future model releases,” the company added.

    SpaceXAI and Meta add to a crowded week of AI launches

    OpenAI’s update comes as other technology companies race to bring new AI models and tools to market.

    Elon Musk’s AI venture SpaceXAI, also known as xAI, is reportedly preparing to release a new AI model with Anysphere, the company behind the AI coding tool Cursor, according to The Information, which cited a memo sent to staff.

    The Information reported that the model could be released as early as Wednesday and is expected to process information quickly, making it competitive in some respects with Anthropic’s Opus 4.8 and OpenAI’s GPT-5.5.

    Meanwhile, Meta launched Muse Image this week, its first image-generation model built by Meta Superintelligence Labs.

    Like other image generators, Muse Image supports prompt-based image generation and editing. Meta says the model can “act as the creative partner” and help users “turn ideas into high-quality visuals” to share on their feed, story or chat.

    However, the model has drawn criticism because users with public Instagram accounts can be mentioned in prompts, allowing others to generate images that use their public posts as reference material, unless they opt out in settings.

    Meta’s policy states that “people may be able to create content with your Instagram content using AI features at Meta” and that users “will not be notified about content created using AI features at Meta.”

    Go to accessibility shortcuts

    Share

    Comments

    Add Euronews on Google

    Read more

  • „Google“ pristatė naują DI vaizdų modelį: žada geresnę kokybę ir mažesnes išlaidas

    „Google“ pristatė naują DI vaizdų modelį: žada geresnę kokybę ir mažesnes išlaidas

    Dirbtinis intelektas vis sparčiau keičia vaizdų ir vaizdo įrašų kūrimą, o „Google“ plečia savo „Gemini“ ekosistemą naujais įrankiais kūrėjams ir verslui. Bendrovė paskelbė apie naują vaizdams skirtą modelį „Nano Banana 2 Lite“ ir išplečia bandomąją „Gemini Omni Flash“ prieigą, kuri orientuota į vaizdo generavimą.

    „Google“ „Nano Banana 2 Lite“ pristato kaip greitą ir ekonomišką „Gemini Image“ šeimos sprendimą, skirtą darbui beveik realiuoju laiku. Jis taikomas scenarijams, kur svarbus mažas vėlavimas ir didelės užklausų apimtys, pavyzdžiui, masinei grafikos generacijai ar automatizuotai produktų vizualizacijai.

    Naujasis modelis siūlomas kaip alternatyva ankstesniam „Nano Banana“ variantui, o atnaujinimas turėtų reikšti geresnę vaizdo kokybę ir greitesnį generavimą. „Google“ taip pat akcentuoja mažesnes eksploatacines sąnaudas, kas aktualu įmonėms, kurios DI turinį generuoja dideliais kiekiais.

    Kur jis jau pasiekiamas?

    „Nano Banana 2 Lite“ jau integruojamas į „Google AI Studio“, „Gemini“ API ir „Gemini Enterprise Agent Platform“. Bendrovė nurodo, kad sprendimas numatytas ir platesniam „Gemini“ produktų rinkiniui, kuriame DI pasitelkiamas kūrybiniam turiniui kurti bei redaguoti.

    Toks plėtimas atspindi bendrą rinkos kryptį: DI generuojamas turinys vis dažniau keliasi iš eksperimentinių įrankių į kasdienius produktus. Verslui tai reiškia spartesnį kūrybinį ciklą, o kūrėjams – galimybę automatizuoti dalį rutininės grafikos gamybos, nors kokybės kontrolė ir autorinių teisių klausimai išlieka aktualūs.

    „Gemini Omni Flash“: vaizdas iki 10 sekundžių

    Antra naujiena – bandomoji „Gemini Omni Flash“ versija, pristatyta „Google I/O“ konferencijos kontekste. Šis sprendimas sujungia multimodalinį „Gemini“ supratimą su natyviu vaizdo įrašų generavimu ir redagavimu, kai įvestimi gali būti tekstas, vaizdai ar vaizdo duomenys.

    Bandomojoje versijoje taikomi ribojimai: generuojami klipai negali viršyti 10 sekundžių, taip pat nurodomi apribojimai darbui su garso nuorodomis ir tam tikromis API funkcijomis. Tokie ribojimai įprasti bandomiesiems leidimams, kai funkcijos palaipsniui plečiamos, o infrastruktūra testuojama realiomis apkrovomis.

    Kainodara pateikiama pagal sugeneruotą trukmę: 10 JAV centų už sekundę, tai yra apie 0,09 euro už sekundę. Praktikoje tai reikštų, kad 10 sekundžių klipas kainuotų maždaug 0,90 euro, tačiau galutinė kaina gali priklausyti nuo planų, naudojimo apimties ir papildomų paslaugų.

    Turinio žymėjimas ir patikimumas

    „Google“ taip pat pabrėžia suderinamumą su „SynthID“ – technologija, skirta DI sugeneruoto turinio žymėjimui. Tokie sprendimai vis dažniau minimi kaip atsakas į augančią dezinformacijos, klastočių ir autorystės nustatymo problemą, ypač plintant realistiškiems vaizdams ir trumpiems vaizdo klipams.

    Bendrovė nurodo, kad „Gemini Omni Flash“ galima derinti su „Nano Banana 2 Lite“, kad būtų sklandesni darbo srautai tarp vaizdo ir vaizdų generavimo. Tai atitinka tendenciją kurti vieną ekosistemą, kurioje skirtingi DI modeliai papildo vienas kitą ir leidžia greičiau pereiti nuo idėjos prie paruošto turinio.