This July I Was Fired from Simple AI (A Deeply YC Company)
15 điểm · 1 bình luận trên Hacker News.
AI, engineering và paper mới. Chỉ lấy tiêu đề, mô tả ngắn rồi dẫn về bài gốc.
4/4 nguồn đang hoạt động15 điểm · 1 bình luận trên Hacker News.
4 điểm · 0 bình luận trên Hacker News.
12 điểm · 1 bình luận trên Hacker News.
25 điểm · 5 bình luận trên Hacker News.
31 điểm · 12 bình luận trên Hacker News.
107 điểm · 54 bình luận trên Hacker News.
13 điểm · 3 bình luận trên Hacker News.
26 điểm · 10 bình luận trên Hacker News.
30 điểm · 8 bình luận trên Hacker News.
31 điểm · 16 bình luận trên Hacker News.
207 điểm · 137 bình luận trên Hacker News.
108 điểm · 6 bình luận trên Hacker News.
183 điểm · 185 bình luận trên Hacker News.
88 điểm · 17 bình luận trên Hacker News.
156 điểm · 87 bình luận trên Hacker News.
19 điểm · 2 bình luận trên Hacker News.
77 điểm · 60 bình luận trên Hacker News.
262 điểm · 158 bình luận trên Hacker News.
268 điểm · 157 bình luận trên Hacker News.
227 điểm · 45 bình luận trên Hacker News.
189 điểm · 114 bình luận trên Hacker News.
136 điểm · 50 bình luận trên Hacker News.
184 điểm · 18 bình luận trên Hacker News.
217 điểm · 173 bình luận trên Hacker News.
192 điểm · 261 bình luận trên Hacker News.
166 điểm · 133 bình luận trên Hacker News.
422 điểm · 321 bình luận trên Hacker News.
257 điểm · 65 bình luận trên Hacker News.
332 điểm · 140 bình luận trên Hacker News.
Signing the code reinforces our commitment to transparency and responsible AI development in Europe.
104 điểm · 31 bình luận trên Hacker News.
Despite rapid progress, most existing vision-language models (VLMs) built from 2D visual inputs often struggle when handling various 3D tasks that require fine-grained spatial understanding and reasoning. To…
Multi-agent interactive world models should not only generate consistent observations, but also maintain world states that persist across agents and evolve across views. Existing autoregressive video diffusion…
Scene understanding requires simultaneous prediction about geometry, appearance, and semantics. However, existing task-specific annotations are fragmented across incompatible, domain-specific datasets. Current…
Diffusion and flow-matching models dominate conditional image generation, yet inference-time scaling for these models is far less developed than for autoregressive language models. Because final quality is…
Flow-based generative models have enabled remarkable progress in fast and controllable generation across continuous and discrete state spaces, yet existing parameterizations are constrained to fixed dimensions…
Compositional generalization is essential for robot to follow diverse instructions. However, pretrained policies are known to take shortcuts, deferring to salient cues rather than grounding language. We…
Controllable video generation remains challenging due to the difficulty of specifying precise multi-object interactions using text prompts or motion-control inputs that primarily constrain pixel movement. In…
Barzilai--Borwein (BB) method has shown strong practical performance in continuous optimization, yet its convergence dynamics remains poorly understood. In particular, a central unresolved question is whether…
Quality control in printing, particularly in rotogravure printing, still depends on slow, costly, and subjective manual inspection. Automated surface defect detection is critical for maintaining high-quality…
Understanding motion in video is a fundamental challenge for visual learning, as frame-to-frame change entangles two sources of dynamics: camera motion and object motion. This decomposition has remained…
Surprisal theory holds that the human processing difficulty of a linguistic unit in context is an affine function of its surprisal under some language model. I argue this claim is a tautology without further…
Faithful explanations of time-series classifiers should identify subsequences that are not only sufficient to preserve a black-box model's prediction, but also necessary for maintaining it. However, existing…
Large Language Models (LLMs) show promise for medical education, but most existing systems focus on localized interactions such as question answering or single-turn feedback, rather than organizing an entire…
Gradient-based saliency methods reveal which input features most influence a neural network's output, and are a standard tool for model interpretability. We observe that differentiable renderers, which are…
Molecular property prediction from structure often uses a single representative conformation, even though many molecules exist as conformational ensembles in solution. We introduce EnsembleEGNN, a molecular…
A consensus anomaly detection framework was applied to monthly malaria surveillance data from Ghana (2014-2023) to identify atypical transmission patterns. Anomalies were highly structured in space and time.…
Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophancy as a one-dimensional failure mode. Models must…
Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to external systems. While powerful, these complex harnesses…
On-policy self-distillation (OPSD) is promising as it removes the external teacher required by on-policy distillation (OPD), yet it still needs asymmetric information between teacher and student to ensure that…
We introduce SANA-Video 2.0, a hybrid video diffusion transformer instantiated at 5B and 14B scales under a unified architecture. Designed to generate high-quality video up to 720p on a single GPU, SANA-Video…
Unlike large language models (LLMs) that exhibit strong reasoning capabilities, vision-language models (VLMs) struggle with visual reasoning, even on geometry problems that admit equivalent text, diagram, and…
While large audio-language models have achieved remarkable progress in auditory perception, they still lag behind text-based large language models in deep logical reasoning, primarily due to the scarcity of…
The coupled ghost and gluon Dyson--Schwinger equations (DSEs) of four-dimensional Landau-gauge Yang--Mills (YM) theory are solved with a neural representation trained only from renormalized equation residuals.…
The rapid progress of AI has intensified the long-standing pursuit of automation: replacing human participation with algorithms wherever possible. Implicit in this pursuit is the assumption that humans remain…
47 điểm · 29 bình luận trên Hacker News.
Learn more about Google for Startups Gemini Startup Forum and apply by August 28.
92 điểm · 10 bình luận trên Hacker News.
We’re releasing the first Activity, Task, Landscape, and Adoption Study (ATLAS) report, showing how people use Google’s AI tools.
Selfie video gives you more options if you’re ever locked out or don’t have access to your usual phone or computer.
Health in ChatGPT now lets eligible U.S. users securely connect medical records and Apple Health to get more personalized insights and better understand their health.
7 điểm · 2 bình luận trên Hacker News.
Have you ever struggled to describe something you’re looking at? Whether it’s a complex manual, a blinking error code, or a unique object, sometimes you just need an exp…
OpenAI announces Project Camellia in Effingham County, Georgia, with commitments to responsible energy, community investment, jobs, and access to Codex.
News organizations are using AI to strengthen reporting, grow audiences, and improve business operations, with OpenAI tools supporting journalists and publishers worldwide.
We shared how Samsung users can boost productivity and get time back on new foldables, watches, and glasses coming soon.
67 điểm · 9 bình luận trên Hacker News.
OpenAI outlines its commitment to advancing American science working with the U.S. Department of Energy and national labs to use frontier AI to accelerate discovery.
50 điểm · 12 bình luận trên Hacker News.
Introducing OpenAI Presence, a proven enterprise AI agent platform that helps organizations deploy trusted voice and chat agents for customer and internal workflows.
NTT DATA Group uses ChatGPT Enterprise and Codex to help 9,000 employees automate work, cut incident analysis to 30 minutes, and scale secure AI adoption.
Discover how college students can use Google AI tools, like Gemini and AI Mode, to prep for grad school, internships and fall courses.
12 điểm · 0 bình luận trên Hacker News.
OpenAI launches the ChatGPT for Small Businesses program, helping entrepreneurs build AI skills, automate work, and grow with ChatGPT Work.
We’re introducing new Gemini models, including Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber.
OpenAI and Hugging Face share early findings from a security incident during AI model evaluation, highlighting advanced cyber capabilities and lessons for defenders.
David Vélez and Robin Vince join the boards of the OpenAI Foundation and OpenAI Group PBC, bringing global leadership in finance, technology, and governance.
Launching a side business? Use Gemini to design your side hustle, conduct market research and automate logistics.
OpenAI shares lessons from deploying long-running AI models, highlighting new safety risks, observed failures, and improved safeguards through iterative deployment.
Sarah Friar, CFO of OpenAI, introduces a practical AI scorecard to measure ROI through useful work, cost per successful task, dependability, and return on compute.