HACKOBAR_
ABOUT
AI News Archive
150 items · 50 per page
1
Local LLM RPG Uses Deterministic State Logic for NPC Emulation
9m ago
0.16
2
Context Blindness in Legal AI Reviews Leads to Organizational Friction
1h ago
0.11
3
CAGE: Enhancing RAG via Coherence-Aware Graph Encoding
1h ago
0.26
4
Fidelity-Aware Training for Coding Agents via C-DPPO
1h ago
0.30
5
Pitch-class Steering for Stable Audio Open via Latent-space Probes
1h ago
0.23
6
SQL-Zero: Zero-Shot Self-Evolving Text-to-SQL via Proposer-Solver Self-Play
1h ago
0.33
7
Cost-Aware Agent Architecture Achieves 91.7% Correctness on DevRev NL2SQL
2h ago
0.37
8
RAAI Framework Optimizes Multilingual Reasoning via Adaptive Inference
2h ago
0.26
9
RCBNB-MB Algorithm Discovers Causal Structures in Non-Stationary Time Series
2h ago
0.20
10
ERPBench Evaluates LLM Agents in Competitive Market Simulations
2h ago
0.23
11
HackProbe Detects and Immunizes Reward Hacking in Self-Evolving Models
2h ago
0.33
12
Week 6 of making my fishing game entirely with AI
2h ago
0.14
13
Qlippy: A Retrieval-Augmented GenAI Assistant for Reproducible Quantum Workflows and Experiment Tracking
2h ago
0.17
14
Rhythms of Work: Multi-Scale Interpretation of Human Behavioral Traces for Workplace Agents
2h ago
0.17
15
SiLR: Structure-Preserving Admission and Process Reward for LLM Tool Agents
2h ago
0.17
16
Extremely Sparse Supervision Incentivizes Reasoning Ability
2h ago
0.17
17
Tracing Audio Grounding and Answer Selection in Audio LLMs
2h ago
0.17
18
Debate Over Open-Weight Model Safety and Economic Viability
3h ago
0.12
19
MCPO Compresses Multimodal Chain-of-Thought via Modality-Contrastive Optimization
3h ago
0.26
20
Scale-QLoRA Enables Efficient Merging for 4-bit Microscaling LLMs
3h ago
0.30
21
Identity Handoff Failures in Grounded Language Model Pipelines
3h ago
0.20
22
ConfRAG Reduces Hallucinations to Below 5% via Confidence Training
3h ago
0.34
23
Internal Probes Reveal Knowledge Hidden by Saturated Decision Thresholds
3h ago
0.23
24
Falcons AI ViT NSFW Classifier Reaches 50M Monthly Downloads
3h ago
0.24
25
Reducing Whisper Hallucinations by 92% via Hallucination Space Projection
3h ago
0.38
26
Interpretable YOLOv10 Detection Using Kolmogorov-Arnold Networks and BLIP
3h ago
0.20
27
La Agente Optima Framework Automates Bayesian Optimization in Self-Driving Labs
3h ago
0.23
28
Octopus Protocol Enables One-Shot Hardware Control via Infrastructure-as-Prompts
3h ago
0.34
29
Indirect Prompt Injection Formulated as a Test-Time Search Problem
4h ago
0.30
30
Comparables XAI Uses Counterfactual Trace Adjustments for Faithful Explanations
4h ago
0.17
31
Atlas Optimizes Compound AI Workflow Deployment on Heterogeneous Clusters
4h ago
0.20
32
Calibrated Reflection Framework Enhances LLM Confidence Estimation
4h ago
0.23
33
TradingAgents v0.4.0 Adds GPT-5.6 Support and Point-in-Time Fixes
4h ago
0.36
34
X-VC Enables Zero-Shot Streaming Voice Conversion
4h ago
0.30
35
Masked Boundary Pause Tokens Improve LLM Reasoning
4h ago
0.34
36
Escalate Framework Enables Controlled Boundary-Aware Safety Refusal
4h ago
0.30
37
Targeted LoRA Fine-Tuning Mitigates LLM Cultural Misalignment
4h ago
0.27
38
KnowChange Uses VLMs for Knowledge-Guided Remote Sensing Synthesis
4h ago
0.42
39
FactoSR Factorizes 4D Properties for Improved Spatial Reasoning
4h ago
0.55
40
Evaluating Test-Time Compute for Conversational Artifact Revision
4h ago
0.48
41
Teacher-Gated On-Policy Distillation Prevents Misleading Updates
4h ago
0.69
42
FlowBalance Calibrates Self-Improvement Using Verifier-Grounded Advantage
4h ago
0.76
43
7 days of making a cozy game with no dev experience. Still no name but I made a cute trailer
5h ago
0.17
44
When LLM Decompilers Recompile More and Preserve Less
5h ago
0.18
45
AI-Powered CPS-Enabled Vulnerable-User-Aware Urban Transportation Digital Twin: Methods and Applications
5h ago
0.18
46
PerfReasoning: How Well Do LLMs Reason on Hardware Performance?
5h ago
0.18
47
A Systematic Evaluation of Cross-Lingual Consistency Enhancement Methods in Multilingual Language Models
5h ago
0.18
48
Shared circuits predict whether LLMs generalize across formats in arithmetic reasoning
5h ago
0.18
49
The VMs Powering Mobile Agents (Instinct, Claude Code)
5h ago
0.24
50
Led by @Samsung, co-led by @eqt's Scaleup Europe Fund and @PSG_equity, with continued backing from @ASMLcompany, @nvidia, @BNPParibas CIB and other…
5h ago
1.08
page 1 / 3
next →