How interest changes
History starts here
The chart will appear after repeat observations. The current metric comes from the source.
23 source pointsReal observations only. History before source connection is not reconstructed.
The text mentions that Btrfs, ZFS, and bcachefs file systems are being tested under workloads, but specific classic benchmarks are skipped.
The chart will appear after repeat observations. The current metric comes from the source.
23 source pointsReal observations only. History before source connection is not reconstructed.
LimiX-2 is a new model in the LimiX family that uses the Contextual Mechanism Networks (CMNs) paradigm, pretrained with Context-Conditional Masked Modeling (CCMM). It focuses on learning joint representations of data generation processes rather than target-centric prediction, showing superior performance on tabular data benchmarks and promoting causal awareness.
4 days agoDaily PapersThis paper evaluates the ability of the MiniMax-H3 multimodal generative model to reason about the physical world. The study introduces a new framework for assessing physical reasoning through four dimensions, using tasks that require integrating information across multiple modalities. The model achieves an overall success rate of 41.97% across 517 evaluation instances, with video-based reasoning performing best and audio-based reasoning least effectively.
3 days agoDaily PapersPhysBrain 1.5 is a unified model for understanding physical environments, generating actions, and predicting future states. It uses a vision-language foundation, encodes language responses and motion as sequences, and is pre-trained on human interaction videos. The model achieves strong performance on embodied understanding benchmarks.
5 days agoDaily PapersThe paper introduces XConf, a method for estimating confidence in language models by leveraging the model's accumulated experience. XConf uses past episodes, including tasks, reflections, confidence levels, outcomes, and lessons learned, to inform current confidence estimates through a recall and reflect process. It outperforms existing methods in discrimination and calibration with lower computational cost.
4 days agoDaily PapersLynnReal-Omni is a native multimodal video generation framework that combines text-to-video, image-conditioned generation, and structural control into a single model. It includes a 32B shared multimodal diffusion transformer and a 27B Flash version for real-time rendering, with a data pipeline and evaluation method designed for video generation tasks.
5 days agoDaily PapersThe paper introduces ScienceIDE, a framework that transforms scientific code repositories into programmable environments for training scientific agents. These environments enable task generation, execution, and verification, supporting supervised fine-tuning and reinforcement learning. The approach demonstrates improvements in scientific code repair and general-purpose benchmarks.
3 days ago