Measuring the Creativity Potential of LLM Agents
Trying to answer the question of "Can LLM agents discover?" through the lens of creativity The post Measuring the Creativity Potential of LLM Agents appeared first on Towards Data Science.
Daily edition · AI Research
TILens turns technical updates into a focused daily brief: official releases, trusted reporting, and practitioner analysis, deduplicated and organized by topic.
502 articles · 2 sources · 502 papers ·
Top topics: AI Research · AI · Python
Trying to answer the question of "Can LLM agents discover?" through the lens of creativity The post Measuring the Creativity Potential of LLM Agents appeared first on Towards Data Science.
A from-scratch PyTorch build that recovers blood flow, viscosity, and wall shear stress in a narrowed artery from 40 noisy velocity readings The post How to Use a PINN for a Navier-Stokes Inverse Problem appeared first…
arXiv:2610.00906v1 Announce Type: cross Abstract: Automated harness optimization can substantially improve LLM agents by iteratively updating their prompts, tool interfaces, and control logic from execution feedback.…
arXiv:2610.01238v1 Announce Type: new Abstract: Unified language models are increasingly expected to combine heterogeneous capabilities, such as mathematics, code, instruction following, and controllable thinking…
arXiv:2610.00420v1 Announce Type: cross Abstract: A weight space network (or metanetwork) takes the weights of another neural network as input and predicts properties of it. Most prior work trains such models on input…
arXiv:2610.00895v1 Announce Type: new Abstract: Foundation models remain vulnerable to spurious correlations and ``Clever Hans'' strategies. Explainable machine learning can find and remove such strategies for…
arXiv:2608.12573v2 Announce Type: replace Abstract: Top-k selection is a fundamental computational primitive with applications spanning databases, information retrieval, signal processing, and modern machine learning…
arXiv:2601.23135v2 Announce Type: replace Abstract: Reinforcement learning (RL) has become a key driver of language model reasoning. Among RL algorithms, Group Relative Policy Optimization (GRPO) is the de facto…
arXiv:2610.02195v1 Announce Type: new Abstract: The generalized Schr\"odinger bridge on a graph moves mass between two distributions while charging a cost for the states visited. It has been approached by learning the…
arXiv:2610.01399v1 Announce Type: new Abstract: Among on-policy deep reinforcement learning methods, Proximal Policy Optimization (PPO) has become the de facto standard, due to its consistently strong empirical…
arXiv:2610.00328v1 Announce Type: cross Abstract: Structured tool calls often fail after only a small number of fields violate a schema or an execution contract. Regenerating the complete object enlarges the action…
arXiv:2610.01369v1 Announce Type: new Abstract: Understanding a nonlinear dynamical system from time series requires not only reproducing its trajectories, but also identifying a simple representation that preserves its…
arXiv:2610.00377v1 Announce Type: new Abstract: Station-based weather forecasting supports daily life and economic activity, yet accurate forecasts require modeling complex spatial dependencies among stations. Recent…
arXiv:2410.10464v3 Announce Type: replace Abstract: Graphs are a highly expressive abstraction for modeling entities and their relations, such as molecular structures, social networks, and traffic networks. Deep Graph…
arXiv:2610.00279v1 Announce Type: cross Abstract: The segmentation of anatomical structures in medical images and particularly in MRI scans, is essential for clinical diagnosis and monitoring disease progression. While…
arXiv:2610.00385v1 Announce Type: new Abstract: Replay selectors often rank cached trajectories by format feedback, confidence, freshness, or response length, although cache-level correctness and downstream learner…
arXiv:2610.00380v1 Announce Type: new Abstract: Depression is a mental condition that can lead to suicide and self-harm. Predicting the outcome of depression treatment is one of the most difficult tasks for clinicians.…
arXiv:2610.01846v1 Announce Type: cross Abstract: Speech-based Alzheimer's disease (AD) assessments increasingly rely on pretrained self-supervised learning (SSL) models that learn acoustic representations directly from…
arXiv:2610.00188v1 Announce Type: new Abstract: Voxel-based volumetric mapping is fundamental to 3D reconstruction, yet fixed-resolution grids remain inherently inefficient - wasting memory in uniform regions and losing…
arXiv:2610.00649v1 Announce Type: cross Abstract: Synthetic speech detection is critical for audio security, but performance can degrade when labeled data are scarce and evaluation conditions differ from training. This…
Showing 1 day · 502 items available