Tootfinder

Opt-in global Mastodon full text search. Join the index!

@Mediagazer@mstdn.social
2026-06-11 20:15:51

Lionsgate and gen-AI company Runway expand their partnership, giving Lionsgate an equity stake in Runway and launching a program to develop and produce new IP (Corbin Bolies/Variety)
variety.com/2026/film/news/lio

@Techmeme@techhub.social
2026-08-13 02:05:42

Sources: Anthropic is in talks to buy Decart, which offers real-time generative video and GPU optimization tech, for about $6B (Bloomberg)
bloomberg.com/news/articles/20

@newsie@darktundra.xyz
2026-08-12 18:01:39

Twitch is Mining Peoples' Streams to Train Amazon's AI 404media.co/twitch-training-am

@sherold@mastodon.online
2026-07-10 12:53:42

An interesting take on the series “Pluribus” by Indrapramit Das. As someone living in Europe, it has made me realize which layers of the story I had missed due to cultural biases.
reactormag.com/resisting-the-h

@ddrake@mathstodon.xyz
2026-06-11 19:02:31

CS education folks!
I just read a very interesting paper:
Margaret-Anne Storey. 2026. From Technical Debt to Cognitive and Intent Debt: Rethinking Software Health in the Age of AI. doi.org/10.48550/arXiv.2603.22
It gives us a couple good ways to describe to studen…

@ErikUden@mastodon.de
2026-06-11 13:11:12

Ever since generative AI, these startups have realized that they can use fear and public outcry over their product as free marketing. Every time a Sam Altman type talks about the “singularity” and how “dangerous” their AI is, they rise in popularity.
It's a horrible marketing discovery, as these concepts are in fact dangerous and morally problematic, so you're caught giving them publicity because the better alternative can't be to not talk about it.

@arXiv_csHC_bot@mastoxiv.page
2026-08-12 08:20:44

Narrative Keyframing for Generative Creative Writing
Chao Zhang, Abe Davis
arxiv.org/abs/2608.10337 arxiv.org/pdf/2608.10337 arxiv.org/html/2608.10337
arXiv:2608.10337v1 Announce Type: new
Abstract: We introduce narrative keyframing, an interaction technique for AI-assisted creative writing that lets writers specify different types of narrative constraints at selected moments in a story, then use AI to generate intervening prose. Inspired by the use of keyframing in animation, narrative keyframing offers a flexible way to connect story planning with adaptive control over generated text. We explore three types of keyframes: plot keyframes define significant events in a story, character keyframes represent how individual characters change over the narrative, and perspective keyframes capture how individual characters experience different events through first-person narratives. Plot and character keyframes offer a flexible way to adapt the type of high-level conditioning explored in previous AI writing tools to more customizable, iterative, and fine-scale control, while perspective keyframes add a new way to control characterization and focalization by using first-person narratives as an intermediary. Through a user study, we show that narrative keyframing supports a more controllable, transparent, and engaging way to use generative AI in creative writing.
toXiv_bot_toot

@jkohlmann@mastodon.social
2026-06-10 02:38:57

It sucks that Apple is undeniably wielding these deep integrations between generative AI, Apple operating systems, and customer data as a cudgel against the EU. apple.com/newsroom/2026/06/due

@Techmeme@techhub.social
2026-07-08 13:21:47

Kaon AI, which builds personalized story worlds using its AI-based FlowGPT and Emochi tools, raised $60M from B Capital and others, and says Emochi has 2M DAUs (Corbin Bolies/Variety)
variety.com/2026/digital/news/

@Mediagazer@mstdn.social
2026-07-08 16:25:42

Kaon AI, which builds personalized story worlds using its AI-based FlowGPT and Emochi tools, raised $60M from B Capital and others, and says Emochi has 2M DAUs (Corbin Bolies/Variety)
variety.com/2026/digital/news/

@gwire@mastodon.social
2026-06-08 16:43:11

I'd missed the "Whata Bod, Idaho" story in Jan, where the US National Weather Service used gen AI and it make up town names.
oecd.ai/en/incidents/2026-01-0

@mxp@mastodon.acm.org
2026-07-07 19:27:17

“When the time comes, the AI industry must burn. It must be allowed to die. Generative AI has already been given far too much money, oxygen and attention, and if it cannot survive without continual venture capital and media coddling, it is unworthy and unnecessary, and must face the cold, hard reality that every regular person faces when they fail.”
Amen to that!

@Techmeme@techhub.social
2026-08-12 22:35:52

Twitch says it intends to use videos streamed on its platform to help train Amazon's generative AI models and tells creators how to opt out (Amanda Silberling/TechCrunch)
techcrunch.com/2026/08/12/amaz

@arXiv_csHC_bot@mastoxiv.page
2026-08-12 08:29:20

AI-Generated Interactive Fiction for Educational Use: A Pilot Study of Perceived Comprehensibility, Coherence, and Engagement
Finn Rogosch, Andreas Schrader
arxiv.org/abs/2608.10818 arxiv.org/pdf/2608.10818 arxiv.org/html/2608.10818
arXiv:2608.10818v1 Announce Type: new
Abstract: Generative artificial intelligence (AI) can produce educational content at scale, including interactive and narrative learning experiences, but technical generation alone is not sufficient: scenarios that are confusing, narratively inconsistent, or unengaging are unlikely to be useful in practice. This paper presents a pilot user-centred evaluation of AI-generated interactive fiction (IF) for educational use in higher education. Using a previously described domain-agnostic pipeline and a shared STEM content base, we generated a controlled pool of scenarios and asked participants (N = 22, STEM higher-education) to play one generated episode and rate it on narrative clarity, story-content coherence, engagement, and length acceptance. A free-text prompt captured open feedback. Narrative clarity and length acceptance were rated positively, engagement sat near the neutral mid-point of the scale, and story-content coherence was the weakest dimension by a clear margin. Qualitative feedback points to quiz integration as the bottleneck. Artificial in-fiction motivation for quiz prompts and abrupt setting changes were reported. Feedback also pointed to missing story-level consequences for wrong answers. From these observations, we derive concrete design implications that can inform larger follow-up studies, including later work on learning effectiveness.
toXiv_bot_toot

@Mediagazer@mstdn.social
2026-08-12 22:15:42

Twitch says it intends to use videos streamed on its platform to help train Amazon's generative AI models and tells creators how to opt out (Amanda Silberling/TechCrunch)
techcrunch.com/2026/08/12/amaz

@ErikJonker@mastodon.social
2026-07-31 11:25:58

Al wat ouder (2024) maar fijn paper over AI & Software engineering/coding. Zowel voor "believers" als totale sceptici 🙂
"tl;dr: Chill, y'all: AI Will Not Devour SE"
arxiv.org/abs/2409.00764

How consequences and human oversight affect the level of validation needed .
From, https://arxiv.org/abs/2409.00764
@privacity@social.linux.pizza
2026-08-05 12:53:23

FPF and Leading Companies Release Risk Assessment Framework and Updated Best Practices for AI in Hiring & Employment
fpf.org/press-releases/fpf-and

@mxp@mastodon.acm.org‬
2026-07-07 19:27:17

“When the time comes, the AI industry must burn. It must be allowed to die. Generative AI has already been given far too much money, oxygen and attention, and if it cannot survive without continual venture capital and media coddling, it is unworthy and unnecessary, and must face the cold, hard reality that every regular person faces when they fail.”
Amen to that!

‪@mxp@mastodon.acm.org‬
2026-07-07 19:27:17

“When the time comes, the AI industry must burn. It must be allowed to die. Generative AI has already been given far too much money, oxygen and attention, and if it cannot survive without continual venture capital and media coddling, it is unworthy and unnecessary, and must face the cold, hard reality that every regular person faces when they fail.”
Amen to that!

@v_i_o_l_a@openbiblio.social
2026-06-03 12:31:39

"Articulating generative AI information literacy competencies: An ACRL Framework–driven model for academic libraries"
doi.org/10.11645/20.1.883

@arXiv_csHC_bot@mastoxiv.page
2026-08-12 08:22:14

MazzikaAI: A knowledge-based performance-to-prompt compiler for real-time Arabic maqam accompaniment with a streaming text-to-music model
Jiaxin Du, Boulbaba Abdeljaouad, Yong Zhuang, Haoyu Li
arxiv.org/abs/2608.10360 arxiv.org/pdf/2608.10360 arxiv.org/html/2608.10360
arXiv:2608.10360v1 Announce Type: new
Abstract: Arabic maqam music microtonal, modal, and built on ornamented call and response is among the traditions most underserved by generative music models, whose training frameworks remain predominantly Western and equaltempered. Real time accompaniment sharpens this gap: an AI partner must listen, adapt dynamically, and respect idiomatic microtonal structures. Streaming text to music models provide strong generative capabilities but lack precise control interfaces. We present MazzikaAI, a knowledge based system that uses natural language as the actuator of a realtime control loop. By compiling live MIDI, gesture, and inferred harmony into continuously updated text prompts, MazzikaAI steers an unmodified streaming generator, Google Lyria RealTime, without requiring model finetuning. The system embeds expert knowledge of six core maqamat, characteristic ornaments, and ensemble dynamics, maintaining realtime responsiveness with subsecond keytoaudibleupdate latency. Empirical evaluations demonstrate that dynamic prompt compilation reliably grounds generation in microtonal scales, significantly increasing offgrid quartertone content over baseline generation. Beyond its core implementation, MazzikaAI illustrates how deterministic knowledgebased rules can effectively bridge expert, nonWestern musical traditions and unfinetuned foundation models. This architecture establishes a scalable paradigm for realtime humanAI cocreation, offering a generalizable blueprint for interactive accompaniment, adaptive music education, and culturally inclusive generative audio across diverse global idioms.
toXiv_bot_toot

@aardrian@toot.cafe
2026-06-08 20:54:49

Updates:
• Two more environmental impact bullets: adrianroselli.com/2024/04/what

@elduvelle@neuromatch.social
2026-08-02 21:06:47

Another reason not to publish in #Frontiers... they use genAI to create alt-text for their figures 🤦
Seriously, remind me why we are paying journals when they can't even write their own Alt-text?
#AcademicChatter

screenshot of the text of a fronteirs article saying:

"
Generative AI statement

The author(s) declared that Generative AI was not used in the creation of this manuscript.

Any alternative text (alt text) provided alongside figures in this article has been generated by Frontiers with the support of artificial intelligence and reasonable efforts have been made to ensure accuracy, including review by the authors wherever possible. If you identify any issues, please contact us.
"
@barijaona@mastodon.mg
2026-07-16 18:19:42

Generative AI Is an Engineering Disaster - The Atlantic
"The problem with generative AI, in the industry’s own jargon, is that it does not scale."
theatlantic.com/technology/202

@anneroth@systemli.social
2026-07-05 12:21:32

Der Energieverbrauch von Google ist in einem Jahr, von 2024 bis 2025, von 31 auf 43 Terrawattstunden (TWh) gestiegen.
Der Effekt von generativer KI, also all den Chatbots, ChatGPT etc.
Wir sollten das nicht als unveränderbar hinnehmen.
ketanjoshi.co/2026/07…

@tiotasram@kolektiva.social
2026-07-03 00:28:54

Just finished "A Deepness in the Sky" by Vernor Vinge. A fascinating and epic science fiction novel with fascinating ideas about technology and also future societies. As I'm delving into a lot of sci fi looking for social imagination it's a good find in some ways but disappointing in others. The Qeng Ho culture is interesting, but their society is uninspiring; Vinge's rosy view of "trade" is not one I completely share, and their hierarchies and wealth accumulation are less compatible with their freedoms in my imagination than in Vinge's. That said, his perspective on the possibilities of galactic-scale civilization and the idea of the "age of failed dreams" are fascinating, especially right now, and his detailed ideas about programming, AI, automaton, and "Focus" are extremely relevant right now. Ultimately I didn't love some parts of the dénouement, and there's a lot of "great man theory of history" going on, including (only somewhat ameliorated) a focus on men over women. Vinge's damsels in distress have a lot more agency than usual for the role, but with the exception of Victory Sr. and Jr., the damsels are very much in distress, and Victory Sr. gets overshadowed a lot by Underhill.
It definitely helps one expand their imagination of what long-term and large-scale human existence could look like, which is great and no easy feat, and both the technologies that are written in and those written out are extremely convincing. The only thing I didn't find compelling was the transposition of capitalism onto such space and time scales. It's a very standard feature of sci fi from the last few decades, and I'm sure most readers don't question it, but I've become someone who can no longer imagine "capitalism across the stars" without questioning how realistic it is that its self-destructive tendencies could possibly last even a few more centuries, let alone succeed at interstellar travel.
One last extremely fascinating thing: how closely the strengths and weaknesses of Focus track with the strengths and weaknesses of modern generative AI. That, and the way that the "age of failed dreams" idea can help people imagine beyond generative AI in a positive way.
#AmReading #ReadingNow #Bookstodon

@seeingwithsound@mas.to
2026-07-23 07:44:01

University of Texas study, for blind or low vision participants: Living in an AI-filtered reality: Participant recruitment & eligibility screening form docs.google.com/forms/d/e/1FAI

@Carwil@mastodon.online
2026-07-02 12:32:24

Gradually making my way through Dwarkesh Patel's oral history of Generative AI.
Big picture: Many AI designers seem to believe that because their software is smarter than they expected, it will inevitably become smarter than them.



The Scaling Era

Dwarkesh Patel

MORE

An inside view of the Al revolution, from the people and companies making it happen.

How did we build large language models? How do they think, if they think? What will the world look like if we have billions of Als that are as smart as humans, or even smarter?

In a series of in-depth interviews with leading Al researchers and company founders - including Anthropic CEO Dario Amodei, DeepMind cofounder Demis Hassabis, OpenAl cofounder Ilya Sutskever, MI…
@toxi@mastodon.thi.ng
2026-07-31 10:55:38

Reminder: Just as with Age ID and Chat Control legislation justifications, the AI build out, the price gauging of memory, storage (and CPU/GPU/SOCs too), the emerging hardware leasing & subscription models — all this is not just driven by generative AI in the form of chatbots, media generators, automation, job replacements, FOMO stories and other "obvious" reasons repeated ad infinitum by CEOs and in the media...
The only way these industry shifts and largest tech investments ever …

@arXiv_csIT_bot@mastoxiv.page
2026-06-11 07:43:08

Vision-Language-Action Models Meet World Models: Embodied Agentic AI for Low-Altitude Wireless Networks
Feibo Jiang, Li Dong, Lei Mao, Kezhi Wang, Cunhua Pan, Dong In Kim, Naofal Al-Dhahir
arxiv.org/abs/2606.11618 arxiv.org/pdf/2606.11618 arxiv.org/html/2606.11618
arXiv:2606.11618v1 Announce Type: new
Abstract: Low-Altitude Wireless Networks (LAWNs), composed of Unmanned Aerial Vehicles (UAVs) and other aerial platforms, provide integrated perception, communication, and computation services in low-altitude airspace. However, deploying large generative models in this domain faces three major challenges: 1) Limited embodied action mapping; 2) Inadequate physical environment modeling; 3) Insufficient closed-loop optimization. To address these challenges, this study proposes an Embodied Agentic UAV framework. Centered on a Vision-Language-Action (VLA) model as the execution core, the framework establishes an end-to-end embodied decision-making pipeline from multimodal environmental perception to continuous control generation. In addition, a World Model (WM) is introduced to capture the coupling between UAV actions and environmental state evolution, thereby supporting environment prediction, policy verification, and dynamic optimization. Furthermore, memory and reflection mechanisms are incorporated to form an adaptive closed-loop optimization paradigm of decision, execution, evaluation, and update, thereby enhancing the system's autonomous decision-making capability and continual evolution ability in complex dynamic environments. Experimental results validate its effectiveness in enabling robust, predictive, and sustainable autonomous control in LAWNs.
toXiv_bot_toot

The nation’s largest public four-year university may soon be barred from replacing faculty with generative AI
as a bill backed by a union of professors comes nearer to reaching the governor’s desk.
Few examples exist of the California State University’s attempting to replace faculty labor with generative AI tools,
but the faculty union wants to prevent such efforts from ever getting off the ground.
The bill so far has garnered no opposition from lawmakers and may clear…

@v_i_o_l_a@openbiblio.social
2026-06-29 05:52:19

"Generative AI Meets Cataloging Practice: Findings from a Comparative Pilot Study"
doi.org/10.5860/ital.v45i2.174
"This study evaluates the performance of four generative AI models—ChatGPT, DeepSeek, Gemini, and Copilot—in generating descriptive metadata for …

@pkraus@berlin.social
2026-07-11 19:47:46

Yesterday, I was one of four panelists at the annual Emmy Noether Treffen organised by the @… , talking about how we use (generative) #ai in #research.
Here's the thing: I don't.
The mood was surprisingly skeptical with none of my peers being particularly optimistic, highlighting issues with applying #llm to indigenous studies (Walther Maradiegue), dealing with fabricated bibliographies (Daria Elagina), or possibilities of quick fixes to the underlying tech (Michael Roth).
1/3

@publicvoit@graz.social
2026-07-23 16:20:43

Members of #Codeberg decided to ban repositories that uses #ai:
"You must not share projects that mostly consist of code written by "generative AI"-tools [...]. Such projects having an unclear copyright status [...] and furthermore have little safeguards to ensure that they do not include h…

@Techmeme@techhub.social
2026-07-23 18:01:13

Runway launches Runway Media Router, which it says is the first built specifically for generative media, as it expands from AI video to AI infrastructure (Rebecca Bellan/TechCrunch)
techcrunch.com/2026/07/23/runw

@ErikJonker@mastodon.social
2026-07-30 12:15:22

There is more then Generative AI / LLMs , nice article about this,
"The important skill is not always choosing the newest model. It is choosing the right model for the problem."
kdnuggets.com/7-machine-learni

@Techmeme@techhub.social
2026-07-10 01:02:00

How the internet, smartphones, social media, and now generative AI are accelerating a drop in the reading of longer works like books (Rose Horowitch/The Atlantic)

@Mediagazer@mstdn.social
2026-07-09 17:20:49

How the internet, smartphones, and now generative AI are accelerating a drop in the reading of longer written works like books, as social media spikes (Rose Horowitch/The Atlantic)
theatlantic.com/magazine/2026/

@tiotasram@kolektiva.social
2026-05-25 18:09:43

Dear generative AI enthusiasts,
Look, I know the tokens you're burning right now don't actually use *that*much energy (even though it's somewhat substantial already and disastrous when we take into account the quality of the crap it's being used for) but what's more important is the appearance (or not) of that token spend on the quarterly earnings report of OpenAI/Anthropic/etc. lays the foundation necessary for those companies to go ahead with their plans for datacenters on a truly ridiculous scale, and those datacenters, if built, ate indeed a climate nightmare which *my kids* will have to live through even if they never benefit from any of it at all. That's (one of many reasons) why I personally need you to stop using generative AI right now.
The fact that the output is crap, the way it erodes your intelligence, and the ways in which it plagiarizes and actively undermines good citation practices are among many other practical reasons not to use it, but what's personal to me is the way that your frivolous sloperation is making the future worse for the baby I'm feeding blueberries to as I type this, and half the time I interact with people like you the conversation begins with some form of "putting aside the ethical issues..."
#AI #GenAI #LLMs

@ripienaar@devco.social
2026-07-30 05:54:43

Somehow didn’t know there was a specific Otel standard for generative AI systems.
Plumber it into my little harness and it’s quite nice.
Here 2 x prompts and a number of tool calls. Need to expand it to instrument a few more things but how nice is that.

@ddrake@mathstodon.xyz
2026-07-20 18:07:55

Reading "Know thine enemy: A critical engagement with AI-assisted software development" by Amy Ko ( medium.com/bits-and-behavior/k

Advanced developers describe a new kind of fatigue with generative artificial intelligence:
"vibecoding fatigue"
or "AI brain fry".
This could affect all language professions, from journalists to lawyers to translators.

@arXiv_eessIV_bot@mastoxiv.page
2026-08-07 07:52:03

A Foundational EDM2-Based Generative Model for High-Resolution Synthetic Fetal Ultrasound Imaging from Open Datasets
Harvey Mannering, Yilin Zhang, Ziao Liu, Zhiwu Huang, Jacqueline Matthew, Miguel Xochicale
arxiv.org/abs/2608.05471 arxiv.org/pdf/2608.05471 arxiv.org/html/2608.05471
arXiv:2608.05471v1 Announce Type: new
Abstract: Prenatal ultrasound imaging is key for assessing fetal health, but AI progress is limited by scarce, privacy-restricted, and hard-to-annotate datasets. We propose a high-resolution fetal ultrasound synthesis framework based on the EDM2 diffusion architecture, trained on multiple public datasets to generate 512x512 images across six anatomical classes. Our method achieved improved image quality with lower FID scores and enhanced downstream fetal plane classification, reaching 93.36% ensemble accuracy after fine-tuning, surpassing real-data-only training. Clinical evaluation by an experienced fetal ultrasound specialist (10 years) on 100 images yielded a mean realism score of 2.67/5, with real images rated higher than synthetic. Artefacts included smoothing, speckle irregularities, and anatomical inconsistencies. Code, data, models and other resources to reproduce this work are available at github.com/xfetus/fetal-ultras.
toXiv_bot_toot

@lapizistik@social.tchncs.de
2026-07-15 09:23:02

Applying the Turing test to the current “chatbot” generative AI does not show that AI is “intelligent”¹ but that humans are very bad in playing the imitation game as judges³.
__
¹In the original paper “ I.—Computing machinery and intelligence”² Turing explicitly states that there is no useful definition for the meaning of the question ‘can machines think?’ and replaces it with: ‘can computers play the imitation game?’
²

@v_i_o_l_a@openbiblio.social
2026-07-02 11:55:23

"How Unique Are Hallucinated Citations Offered by Generative Artificial Intelligence Models?"
doi.org/10.3390/publications14
"This paper investigates how generative AI produces and propagates hallucinated academic references, focusing on the recurring…

@Techmeme@techhub.social
2026-06-18 00:36:04

Noam Shazeer leaves Google to join OpenAI as lead for architecture research; he rejoined Google in 2024 as part of the Character.AI deal and was Gemini co-lead (The Information)
theinformation.com/articles/st

@frankel@mastodon.top
2026-05-14 17:18:12

What I’m Hearing About #CognitiveDebt (So Far)
margaretstorey.com/blog/2026/0

@lpryszcz@genomic.social
2026-06-16 13:34:18

"... we have the right to decide whether and how we want to use technologies. Ideally, this should be in a way that benefits us all...
The crux of the matter is that ethical behaviour does not come for free. Ethics are neither efficient nor do they enhance your economic profit. That means that by acting according to your values you will, at some point, have to give something up. If you’re not willing to do that, you don’t have values - just opinions."

@cjust@infosec.exchange
2026-06-18 17:24:53

No, Artificial Intelligence Is Not Conscious

--Ted Chiang, The Atlantic
Should we seriously consider the possibility that Claude, or any large language model, might be conscious? And if it has feelings, is it capable of receiving moral instruction?
No. Absolutely not. Generative AI is harmful enough when we understand it as a conventional technology, but if we confuse fluency at generating text with consciousness or moral agency, we’re at risk of assigning re…

@Mediagazer@mstdn.social
2026-07-16 21:11:46

Netflix says roughly 300 titles used generative AI this year, mostly in post-production, "to deliver higher quality output more quickly and at a lower cost" (Emma Roth/The Verge)
theverge.com/streaming/966633/

@tiotasram@kolektiva.social
2026-05-26 11:36:22

Are you in tech and outraged about generative AI? Is it being forced down your throat at work?
Here's a nice vindictive way to get a little revenge if you want:
1. Find a project that contains slop code.
2. Optionally, identify specific files or functions that are LLM-generated. I guarantee you that on average, this code has not been adequately tested/inspected, even/especially if it contains LLM-generated test cases.
3. Make up a reason the code could be flawed, bonus points if it's subtle or hard to test. Don't put effort into this or try to actually find a flaw. Just make something up at random.
4. Report your made-up defect as a bug.
That's it. If anyone ever questions you on the incorrect report, just say "oh I used an LLM and it said there was a bug so I reported it." (Don't actually use an LLM, that would be feeding the bubble.)
Note that you are showing the creator of the code the exact same amount of disrespect that they've shown you by publishing slopcode in the first place. I'd bet odds are 50:50 or better that if a human actually follows up on the report, even though they'll find out that the bug report is wrong, they'll find and fix some other subtle flaw in the LLM-generated code, so this is actually helpful in a way.
For step 3, try to get creative. Like "logic in decideUVParameters can cause state to be inconsistent in some cases." If asked for a steps to reproduce, either make one up if it's easy to do so, or say "I forgot how I triggered this." Surely they can ask an LLM to figure out conditions that would trigger the bug ;).
#AI #LLMs #GenAI

@arXiv_csCR_bot@mastoxiv.page
2026-07-24 07:36:29

Deepfake News Detection: A Multimodal Framework Integrating LipNet, DeepSpeech and ResNET for Enhanced Audio-Visual Analysis
Ameena Khan, Muhammad Ahsan Aziz, Muhammad Junaid Asif, Naeem Akhter, Rana Fayyaz Ahmad
arxiv.org/abs/2607.20579 arxiv.org/pdf/2607.20579 arxiv.org/html/2607.20579
arXiv:2607.20579v1 Announce Type: new
Abstract: Deepfake news refers to AI-generated (or AI ma-nipulated) multimedia content intentionally generated to deceive audiences by manipulating the facial expressions, or speech while maintaining the realistic appearance. The rapid progress of generative AI has made the synthesis of highly realistic fake videos and cloned voices widely accessible, posing a serious threat to the authenticity of digital news media. This paper presents a multi-modal framework that discerns the authenticity of video content by jointly exploiting audio and visual cues, thereby addressing the challenge of detecting the deepfake videos. We proposed a framework that involves features extraction from lip movements, audio content and video frames. Lip movements and speech content are encoded using the LipNet and DeepSpeech2 models, while facial features are extracted by leveraging the use of BlazeFace and represented with ResNet18. The extracted feature vectors are concatenated into a holistic video representation and classified with an ensemble of machine learning and deep learning models, including Random Forest (RF), Multi-layer Perceptron (MLP) and Long Short-Term Memory (LSTM) networks. Exten-sive experiments performed on the FakeAVCeleb dataset shows that the proposed approach attains an accuracy of 94% using augmented audio features, outperforming a state-of-the-art multi-modal ensemble baseline. The results confirm the robustness and practical potential of the proposed framework for deepfake news detection.
toXiv_bot_toot

@ErikUden@mastodon.de
2026-06-17 06:28:15

Ever since generative AI began making up large sums of impressions on the internet, advertisers simply don't know anymore what statistics can be believed, hence they are lobbying for age verification laws.

@Techmeme@techhub.social
2026-07-16 21:08:15

Netflix says roughly 300 titles used generative AI this year, mostly in post-production, "to deliver higher quality output more quickly and at a lower cost" (Emma Roth/The Verge)
theverge.com/streaming/966633/

@Techmeme@techhub.social
2026-06-18 21:25:48

Snap plans to spin off an internal generative AI video team into Dotmo, a new company focused on AI models for interactive gaming experiences, citing high costs (Lucas Ropek/TechCrunch)
techcrunch.com/2026/06/18/snap

@Mediagazer@mstdn.social
2026-07-29 03:45:40

Letter: WGAE withdrew its sponsorship of NY film festival Urbanworld after it unveiled a new programming section for directors using generative AI in their work (Brooks Barnes/New York Times)
nytimes.com/2026/07/28/movies/

@mxp@mastodon.acm.org‬
2026-05-17 11:07:26

This article has the merit of showing that academics don’t have to adopt genAI uncritically, but it treats the uncritical adoption as the default, portraying those that take a more reflective approach as “refusing to use generative AI.”
Shouldn’t the burden of proof be on the AI advocates? And shouldn’t academics’ arguments go beyond regurgitating the companies’ advertising and trite claims of “it’s the future, use it or you’ll get left behind”?

@mxp@mastodon.acm.org
2026-05-17 11:07:26

This article has the merit of showing that academics don’t have to adopt genAI uncritically, but it treats the uncritical adoption as the default, portraying those that take a more reflective approach as “refusing to use generative AI.”
Shouldn’t the burden of proof be on the AI advocates? And shouldn’t academics’ arguments go beyond regurgitating the companies’ advertising and trite claims of “it’s the future, use it or you’ll get left behind”?

‪@mxp@mastodon.acm.org‬
2026-05-17 11:07:26

This article has the merit of showing that academics don’t have to adopt genAI uncritically, but it treats the uncritical adoption as the default, portraying those that take a more reflective approach as “refusing to use generative AI.”
Shouldn’t the burden of proof be on the AI advocates? And shouldn’t academics’ arguments go beyond regurgitating the companies’ advertising and trite claims of “it’s the future, use it or you’ll get left behind”?

@Mediagazer@mstdn.social
2026-05-14 16:05:47

Job listings indicate Netflix has been building an internal studio called INKUbator that aims to use AI to produce short-form animated content (Janko Roettgers/The Verge)
theverge.com/column/930118/net

@v_i_o_l_a@openbiblio.social
2026-05-27 12:01:53

"Better Than a Google Search?: Effectiveness of Generative AI Chatbots as Information Seeking Tools in Law, Health Sciences, and Library and Information Sciences"
pal-ojs-tamu.tdl.org/pal/artic
"[…] Using 30 discipline-specific prompts gro…

@Techmeme@techhub.social
2026-07-23 20:05:44

Germany's Black Forest Labs launches Flux 3 and Flux-mimic, its first models for robotics, as it expands from generative AI into physical AI (Yazhou Sun/Bloomberg)
bloomberg.com/news/articles/20

@tiotasram@kolektiva.social
2026-06-29 01:58:24

Thinking about some things that @… said in another thread, and as someone who advocates against AI hype and against the use of most generative AI in most circumstances, I feel it's important to say: many of the ethical issues with using generative AI mirror almost directly the ethical issues with living/working on land stolen by colonists, except that they're less harmful.
Arguments like "well we don't really know whose work it's ripping off this time" and "artists that post their art online know it's going to be looked at; this is the same thing" and "well it's inevitable and everyone's doing it so it's unreasonable to make a big deal about it" directly echo arguments like "well now we don't know whose land it was any more exactly" (yes, we do; you can literally go look up the website of their descendants), or "the natives weren't really using the land anyways", or "it's all in the past now, and it's unavoidable." That unavoidable one is actually somewhat true of using stolen land, at least compared to LLM usage.
If you can see through those lies in the case of AI hype but choose not to do so in the case of colonialism, that says something about your priorities and allegiances.
This is not at all a call for people to talk less about AI; rather it's a call for those who take opposing AI hype seriously to look around and make some noise about other injustices too (I realize many of you already do this).
#AI #LLMs #LandBack #GenAI

@Mediagazer@mstdn.social
2026-07-24 16:40:56

QVC and HSN hosts vote to unionize and join SAG-AFTRA, an organizing drive fueled by concerns around generative AI (Katie Kilkenny/The Hollywood Reporter)
hollywoodreporter.com/business

@v_i_o_l_a@openbiblio.social
2026-05-25 19:45:30

"Teaching citation in the age of generative AI: Rethinking research literacy, academic integrity, and epistemic responsibility"
doi.org/10.1016/j.acalib.2026.

@metacurity@infosec.exchange
2026-07-25 14:31:40

Every week, Metacurity offers our subscribers a rundown of the best infosec-related long reads that we can't do justice to in the crazy crush of daily news.
This week's selection covers
--AI incident reporting laws leave dangerous blind spots,
--Federal data consolidation raises surveillance fears,
--Russia's FSB embraces generative AI,
--Flock camera error triggers police stops,
--Why legacy OT malware won't die
metacurity.com/the-governance-

@thomasfuchs@hachyderm.io
2026-06-26 14:35:26

TIL the Carmina Burana (11th/12th century) had a poem about generative AI use at universities
Florebat olim studium,
nunc vertitur in tedium;
iam scire diu viguit,
sed ludere prevaluit.
iam pueris astutia
contingit ante tempora,
qui per malivolentiam
excludunt sapientiam.
sed retro actis seculis
vix licuit discipulis
tandem nonagenarium
quiescere post studium.
at nunc decennes pueri
decusso iugo liberi
se nunc magistros iactitant,
ceci cecos precipitant,
implumes aves volitant,
brunelli chordas incitant,
boves in aula salitant,
stive precones militant.
Translation:
Once learning flourished. Now it's come
to be condemned as tedium:
the days of thirsting after truth
are now the idle days of youth.
For students hardly in their prime
find themselves wise before their time:
they know it all — impertinence
replaces plain intelligence.
In days gone by we were required
to stick with study: none retired,
or wished himself to be released,
till ninety years of age at least.
Now lads of barely a decade
can graduate — get themselves made
professors too! And who's to mind
how blind the blind who lead the blind?
So fledgelings soar upon the wing,
so donkeys play the lute and sing:
bulls dance about at court like sprites
and ploughboys sally forth as knights.

@Techmeme@techhub.social
2026-05-21 16:05:54

Spotify and UMG plan to let Premium users create AI covers and remixes using music from participating UMG artists as a paid add-on, without giving a launch date (Jem Aswad/Variety)
variety.com/2026/digital/news/

@Mediagazer@mstdn.social
2026-05-21 16:10:50

Spotify and UMG plan to let Premium users create AI covers and remixes using music from participating UMG artists as a paid add-on, without giving a launch date (Jem Aswad/Variety)
variety.com/2026/digital/news/

@Techmeme@techhub.social
2026-07-24 12:40:47

Meta launches Facebook Verified, a free program it says will verify that users are real humans by analyzing a facial recognition selfie and assigning badges (Mat Smith/Engadget)
engadget.com/2222353/meta-laun

@Techmeme@techhub.social
2026-06-17 04:56:45

Sources: Jeff Bezos is investing in Cambridge, UK-Based CuspAI, which applies generative AI to material sciences, as part of a $400M round at a $2.6B valuation (Tim Bradshaw/Financial Times)
ft.com/content/4c479227-567c-4

@Techmeme@techhub.social
2026-06-17 04:56:45

Sources: Jeff Bezos is investing in Cambridge, UK-Based CuspAI, which applies generative AI to material sciences, as part of a $400M round at a $2.6B valuation (Tim Bradshaw/Financial Times)
ft.com/content/4c479227-567c-4

@thomasfuchs@hachyderm.io
2026-06-22 00:32:29

Few people know this but Deep Space Nine is an early example of generative AI use in visual media