Beyond Market Intelligence/generative AI for data analysis

generative AI for data analysis

generative AI for data analysis on Beyond Market Intelligence: a running collection of 207 stories we have gathered and hand-picked because they are worth your time. Every post here touches on generative ai for data analysis in some way — the news, the analysis, the deep dives, and the occasional surprise find. Acme AI is the next-generation, AI-powered spreadsheet platform built to replace Excel and redefine how analysts, data scientists, and enterprise teams work with data. New stories are added to this page as we find them, so check back if you want to keep up with what is happening around generative ai for data analysis, or subscribe to the RSS feed to get them as soon as they are published. Browse the collection below, or head back to the homepage to see everything Beyond Market Intelligence is covering right now.

Meta says Muse Spark 1.3 has frontier performance — but its best results come from a model developers can’t broadly use yet
VentureBeat

Meta says Muse Spark 1.3 has frontier performance — but its best results come from a model developers can’t broadly use yet

Meta's newest AI model, Muse Spark 1.3, delivers notable performance gains over its predecessor, achieving "frontier performance" as CEO Mark Zuckerberg proclaimed. While the most impressive results stem from a "max reasoning" configuration still undergoing safety testing, the broadly available version ranks among the strongest price-performance offerings near the top of independent model evaluations. Though not currently leading the leaderboard—Anthropic’s Claude Fable 5.1 still holds that distinction—Muse Spark 1.3 represents a significant step forward, trading wins with OpenAI and Anthropic on key coding benchmarks.

Microsoft AI’s MAI-Transcribe-2 undercuts OpenAI, Google and ElevenLabs on price and speed
VentureBeat

Microsoft AI’s MAI-Transcribe-2 undercuts OpenAI, Google and ElevenLabs on price and speed

Microsoft AI has significantly disrupted the speech recognition landscape with the release of MAI-Transcribe-2, undercutting OpenAI, Google, and ElevenLabs on both price and speed. Priced at just 10 cents per hour, this represents a remarkable 72% reduction from the initial model's cost. Offering features like speaker diarization, word-level timestamps, and code switching—typically premium capabilities—for this price, MAI-Transcribe-2 positions itself as a compelling solution for enterprises processing substantial audio volumes. For those interested in exploring this evolving market, “Meta prices Muse Voice Transcribe at $0.

We released TontaubeV1, a character-level TTS model for long-form generation [P]
Machine Learning

We released TontaubeV1, a character-level TTS model for long-form generation [P]

We're excited to announce the release of TontaubeV1, a 2.9B-parameter open-weight Text-to-Speech (TTS) model engineered for expressive speech and seamless long-form generation. Primarily supporting English and German, TontaubeV1 leverages innovative character-level tokenization and a unique chunking/position scheme to enhance performance and maintain context even in extended passages. Achieving a 50.1% score on an LLM-as-a-judge audiobook benchmark against ElevenLabs, this model represents a significant advancement in accessible AI-driven voice technology. Explore the model and demo on Hugging Face today.

Anthropic's Claude Fable 5.1 and Mythos 5.1 arrive with a 75% cost reduction for Fable cache reads
VentureBeat

Anthropic's Claude Fable 5.1 and Mythos 5.1 arrive with a 75% cost reduction for Fable cache reads

Anthropic has released Claude Fable 5.1 and Claude Mythos 5.1, the latest iterations of its powerful large language models, alongside a significant 75% cost reduction for Fable cache reads. These models prioritize sustained problem-solving, demonstrating substantial improvements on benchmarks like Terminal-Bench and AutomationBench. Crucially, Anthropic is also introducing Enterprise Frontier Safeguards (EFS), allowing organizations to retain monitoring data within their own infrastructure. This release addresses evolving enterprise needs for capable, economical, and governable AI agents—a shift underscored by recent cybersecurity evaluations.

Your files stay put: Perplexity’s hybrid AI keeps confidential data off the cloud
VentureBeat

Your files stay put: Perplexity’s hybrid AI keeps confidential data off the cloud

Perplexity today introduces hybrid AI compute, a transformative system designed to keep your confidential data secure. Computer, Perplexity’s agentic platform, now intelligently splits tasks between cloud-based and locally-run AI models on Apple silicon Macs, ensuring sensitive information never leaves your device. This innovative approach combines the power of frontier models with the privacy of on-device processing, a critical advancement for industries handling sensitive data. Explore this new capability today and discover how Perplexity is redefining data security and productivity.

AI agents that pass authentication can still drift, expose data, or get memory-poisoned
VentureBeat

AI agents that pass authentication can still drift, expose data, or get memory-poisoned

Securing AI agents requires a shift in perspective. While gateways are often the initial defense, they're frequently deployed before foundational identity and attribution layers are in place, creating a significant vulnerability. Recent events, like the CISA advisory regarding a LiteLLM flaw, highlight this risk. Prioritize establishing agent inventory, distinct identities, and task-scoped credentials *before* relying on runtime enforcement. Start with the basics – identifying and naming your agents – to build a robust security foundation.

I analyzed 31,352 hourly LLM benchmark scores: within-day variation was 2.8 points, while between-day variation was 8.4 [P]
Machine Learning

I analyzed 31,352 hourly LLM benchmark scores: within-day variation was 2.8 points, while between-day variation was 8.4 [P]

A new analysis of 31,352 hourly LLM benchmark scores reveals critical insights into model stability. Examining coding, reasoning, and tool-calling performance, the research found between-day variation (8.4 points) was approximately three times greater than within-day variation (2.8 points), suggesting sustained daily changes offer a stronger signal for detecting performance drift. This work, underpinning the open-source AIStupidLevel system, now encompasses over 169,000 benchmark runs and powers a model router optimizing for performance and cost—a dimension often missing from standard monitoring.

Cohere Parse 5 loses the benchmark on points. It wins on cost per page.
VentureBeat

Cohere Parse 5 loses the benchmark on points. It wins on cost per page.

Enterprises seeking to integrate PDFs, slides, and scanned documents into AI pipelines often encounter a critical bottleneck: balancing accuracy with cost. Cohere’s Parse 5 addresses this challenge, prioritizing price-to-performance over raw accuracy. While benchmark results show Parse 5 trailing larger models like GPT-5.5, it delivers a compelling value proposition, costing just $1.50 per 1,000 pages. This strategic approach makes enterprise-scale document parsing more economical, a crucial step in realizing the potential of agentic AI, as highlighted in our recent article on agentic AI security.

Salesforce just put its entire CRM inside Claude — and says you’ll never need its app again
VentureBeat

Salesforce just put its entire CRM inside Claude — and says you’ll never need its app again

Salesforce and Anthropic are redefining enterprise software with Claudeforce, a new plugin that brings the entire CRM platform directly into Claude. This innovative integration, available to select customers today, empowers sellers to query, update, and act on live CRM data without ever opening Salesforce itself—potentially eliminating thousands of clicks per morning. Salesforce envisions a future where the UI *is* the AI, allowing for dynamic app creation and personalized workflows.

Orchestration is the new challenge for CX in the age of AI agents
VentureBeat

Orchestration is the new challenge for CX in the age of AI agents

The rise of AI agents presents a new challenge for customer experience: orchestration. As enterprises rapidly deploy AI across channels, many are struggling to integrate these tools with legacy systems, creating fragmented customer journeys and overburdened human agents. Tata Communications’ Gaurav Anand explains that the shift is moving away from simple automation toward intelligent orchestration—connecting tasks and delivering end-to-end outcomes with a shared understanding of the customer. Discover how this approach, underpinned by a common enterprise ontology, can transform CX.

Prompt injection ranks No. 1 with OWASP and No. 12 in the incident record. The attack itself is invisible to a scan.
VentureBeat

Prompt injection ranks No. 1 with OWASP and No. 12 in the incident record. The attack itself is invisible to a scan.

Prompt injection currently ranks No. 1 with OWASP, yet real-world incident records place it at No. 12 – a divergence revealing a critical gap in how we assess AI risk. This discrepancy, uncovered by Kyriakos “Rock” Lambros and Steve Wilson, highlights that a low CVE count shouldn’t lull security teams into complacency. While defenses are working, the attack surface remains vast, demanding a shift from reactive vulnerability scanning to proactive architectural controls, like authorization gates, to limit potential damage.

Perplexity partners with Nvidia to launch Portable Computer, a fully local AI agent with zero token costs
VentureBeat

Perplexity partners with Nvidia to launch Portable Computer, a fully local AI agent with zero token costs

Perplexity today launches Portable Computer, a significant step toward bringing powerful AI agents directly to users' hardware. Developed in partnership with Nvidia, this version of Perplexity’s “Computer” platform runs entirely locally, eliminating token costs and prioritizing data privacy. By combining a streamlined agent harness with models like Qwen 3.8, Portable Computer delivers impressive performance, even rivaling frontier models in certain tasks. For those exploring the possibilities of local AI, consider "How to Leverage Local Small Language Models for Your Projects" for a practical guide.

Anthropic’s new Claude Tag update lets its Slack agent read the full conversation — and jump in unprompted
VentureBeat

Anthropic’s new Claude Tag update lets its Slack agent read the full conversation — and jump in unprompted

Anthropic’s latest Claude Tag update marks a pivotal shift in enterprise AI. Now, Claude's Slack agent reads entire conversations, proactively offering assistance—sometimes unprompted—a move Anthropic calls "multiplayer AI." This represents a transition from individual AI tools to collaborative agents embedded within teams, streamlining workflows and boosting productivity. According to Anthropic, this change improves decision-making by roughly 30%.

Machine Learning

Mapping intrinsic rank and informational gravity in complex tabular data: I developed a non-parametric, model-agnostic, information-theoretic diagnostic to bypass the limits of linear, rank, and Euclidean baselines. [R]

Navigating complex tabular data often reveals limitations with standard dimensionality reduction techniques like PCA. To address this, I’ve developed a non-parametric, model-agnostic diagnostic leveraging information theory to bypass these constraints. The "Entropic Scree" accurately maps intrinsic rank and “informational gravity,” distinguishing shared signal from noise and revealing hidden topological structures—even in datasets where features exceed samples. Explore the methodology and open-source framework on GitHub to transform your data exploration and inform architectural decisions for downstream AI models.

One in five enterprises can't stop a runaway AI agent's spending in real time
VentureBeat

One in five enterprises can't stop a runaway AI agent's spending in real time

Enterprise adoption of AI agents is revealing a critical shift: organizations are increasingly deploying multiple orchestration platforms—averaging three—to mitigate vendor risk and retain control. This trend, driven by concerns around security, permissions, and visibility, sees Microsoft AI Foundry/Copilot Studio leading usage, with Anthropic's Claude Platform gaining significant consideration. Notably, one in five enterprises still lacks real-time control over agent spending, highlighting the need for robust oversight as AI deployments evolve. Learn more about this emerging landscape with VentureBeat's coverage of Serval’s AI agent, Catalyst.

VentureBeat names Rob Strechay as its first Lead Analyst, expanding its enterprise AI research push
VentureBeat

VentureBeat names Rob Strechay as its first Lead Analyst, expanding its enterprise AI research push

VentureBeat significantly expands its enterprise AI research capabilities with the appointment of Rob Strechay as its first Lead Analyst. Strechay, formerly of theCUBE Research, brings three decades of experience across practitioner, executive, and analyst roles, uniquely positioning him to address the critical data needs of technical decision-makers. His focus will initially encompass cloud infrastructure, data infrastructure, and AI security, complementing VentureBeat’s VB Pulse surveys—including recent findings on agentic orchestration—to provide objective insights for navigating the evolving AI landscape.

Qwen3.8-27B runs frontier-class coding agents and reasoning locally, no cloud API required
VentureBeat

Qwen3.8-27B runs frontier-class coding agents and reasoning locally, no cloud API required

Alibaba's Qwen3.8-27B model marks a significant shift in the AI landscape, offering frontier-class coding and reasoning capabilities accessible locally—no cloud API required. This 27-billion-parameter model, released under an open-source license, delivers impressive performance, rivaling proprietary models like Claude Opus on key benchmarks. Its compact size, runnable on consumer hardware, empowers developers and enterprises to explore AI-driven solutions with greater privacy, control, and cost-efficiency, fundamentally changing how powerful AI can be deployed.

Enterprises with AI context layers report agent failures at more than twice the rate of those without one
VentureBeat

Enterprises with AI context layers report agent failures at more than twice the rate of those without one

Enterprises are increasingly grappling with a critical challenge: AI agents confidently delivering incorrect answers. Recent data reveals that enterprises with AI context layers report agent failures more than twice as often as those without—a counterintuitive finding highlighting a key truth. Sixty-eight percent trace these errors to inconsistent business context, and the problem is escalating. While adoption of governed context layers is rising, it’s exposing failures rather than preventing them, underscoring the need for robust data governance.

Writer says its new Palmyra X6 model cuts AI agent costs by 52% as token spending surges
VentureBeat

Writer says its new Palmyra X6 model cuts AI agent costs by 52% as token spending surges

Writer today unveiled Palmyra X6, a new AI agent model poised to significantly reduce costs for enterprise users. Paired with its rebuilt agent orchestration “harness,” Palmyra X6 delivers an average of 52% lower operational costs, alongside a 48% speed improvement and 10% quality boost. Leveraging a post-trained version of GLM-5.2, Writer emphasizes control and cost transparency, offering governance tools and multi-model support—a strategy echoing the shift towards pragmatic AI adoption, as explored in "Why Capital One built its multi-agent AI platform around open-weight models."

SpaceXAI debuts Grok 4.6, overtaking Kimi K3's performance and matching GPT-5.6 Sol for world's third best on Artificial Analysis
VentureBeat

SpaceXAI debuts Grok 4.6, overtaking Kimi K3's performance and matching GPT-5.6 Sol for world's third best on Artificial Analysis

SpaceXAI, formerly xAI, has released Grok 4.6, its latest AI model, focused on long-running agents, coding, and knowledge work, offering a competitive pricing strategy. Scoring 61 on the Artificial Analysis Intelligence Index, Grok 4.6 ties OpenAI's GPT-5.6 Sol for the third-best position globally, surpassing Kimi K3. This upgrade delivers significant gains over Grok 4.

Nvidia's Switchyard router reshuffles AI models mid-task, cutting task costs to a third in its own tests
VentureBeat

Nvidia's Switchyard router reshuffles AI models mid-task, cutting task costs to a third in its own tests

Enterprises face a persistent challenge: balancing the power of advanced AI agents with escalating costs. Traditionally, relying solely on frontier models or building custom routing logic proved inefficient. Nvidia proposes a solution with Nemotron 3.5 Lightning, a fast, specialized model, and NeMo Switchyard, an open-source routing library. This pairing delivers frontier-level performance while potentially cutting benchmark costs by a third.

AWS Continuum integrates with OpenAI Codex and Anthropic Claude Code in major AI security push
VentureBeat

AWS Continuum integrates with OpenAI Codex and Anthropic Claude Code in major AI security push

Amazon Web Services is making a significant move to bolster AI security, integrating its Continuum platform—designed to identify code vulnerabilities—directly into coding environments built by OpenAI and Anthropic. This initiative embeds AWS's security tooling where developers write code, regardless of the AI model used. The urgency stems from recent advancements like Anthropic's Claude Mythos Preview, which revealed a surge in previously unknown vulnerabilities, prompting AWS to prioritize autonomous security at machine speed. For deeper insight into AI model capabilities, explore our article on OpenAI’s GPT-5.6-Cyber.

Meta returns to open source with Muse Glimmer, an Apache 2.0 licensed 30B parameter AI model optimized for agents — available now
VentureBeat

Meta returns to open source with Muse Glimmer, an Apache 2.0 licensed 30B parameter AI model optimized for agents — available now

Meta’s return to open source with Muse Glimmer marks a significant shift in the AI landscape. This 30-billion-parameter model, licensed under the permissive Apache 2.0, is specifically optimized for autonomous AI agents and designed to run directly on consumer hardware like Macs and PCs. Unlike previous Meta releases, Glimmer offers unrestricted commercial use and redistribution. The model's ability to operate locally, without cloud dependency, enhances data privacy and reduces costs, as demonstrated by its efficient performance on just 24GB of VRAM.

Tencent's Team Memory shares AI agent memory across a team — with no governance yet for when it's wrong
VentureBeat

Tencent's Team Memory shares AI agent memory across a team — with no governance yet for when it's wrong

Tencent’s Agent Memory, now extended with the beta launch of Team Memory, addresses a critical gap in AI agent technology: enabling teams of agents to leverage a shared context. This open-source project, already trending No. 1 on GitHub, moves beyond individual agent memory, offering a shared hub with reusable assets like Chat Memory, Skill, LLM-Wiki, and Code-Graph.