BTC $83,953.67 +0.88%
ETH $2,717.44 +1.88%
BNB $764.74 +0.55%
XRP $1.55 +4.47%
SOL $120.57 +1.97%
TRX $0.3351 +0.30%
DOGE $0.0956 +3.98%
ADA $0.2531 +4.55%
BCH $311.82 +0.62%
LINK $15.09 +5.83%
HYPE $87.74 +0.57%
AAVE $175.20 +20.38%
SUI $1.17 +0.59%
XLM $0.2319 +7.17%
ZEC $1,442.18 -7.83%
AAPL $332.25 -2.27%
AMZN $246.84 +0.17%
GOOGL $339.80 -0.26%
MSFT $511.62 +1.05%
META $723.95 +0.78%
NVDA $231.20 +1.01%
TSLA $355.02 -1.38%
SNDK $1,737.19 +4.19%
INTC $117.95 +1.95%
SPCX $149.81 +1.91%
MU $1,079.45 +4.35%
AMD $615.88 +2.86%
BTC $83,953.67 +0.88%
ETH $2,717.44 +1.88%
BNB $764.74 +0.55%
XRP $1.55 +4.47%
SOL $120.57 +1.97%
TRX $0.3351 +0.30%
DOGE $0.0956 +3.98%
ADA $0.2531 +4.55%
BCH $311.82 +0.62%
LINK $15.09 +5.83%
HYPE $87.74 +0.57%
AAVE $175.20 +20.38%
SUI $1.17 +0.59%
XLM $0.2319 +7.17%
ZEC $1,442.18 -7.83%
AAPL $332.25 -2.27%
AMZN $246.84 +0.17%
GOOGL $339.80 -0.26%
MSFT $511.62 +1.05%
META $723.95 +0.78%
NVDA $231.20 +1.01%
TSLA $355.02 -1.38%
SNDK $1,737.19 +4.19%
INTC $117.95 +1.95%
SPCX $149.81 +1.91%
MU $1,079.45 +4.35%
AMD $615.88 +2.86%

train

All
Article
Flash

first_img OpenAI: Safety justification should be submitted before cutting-edge reinforcement learning training

On September 28, 2026, OpenAI published a safety-related article, stating that before continuing any cutting-edge reinforcement learning training, structured safety documentation should be required. Ideally, such documentation should reach the evidence-based structured risk argument level used in safety-critical industries like aviation and nuclear power. OpenAI views this as a direction for effort while acknowledging the complexity arising from the emergence of AI capabilities, making it difficult to achieve the same level of rigor.The article focuses on cutting-edge reinforcement learning training and does not cover the broader alignment attributes required for internal and external deployments. The recommendations in the article reflect current practices, which are expected to continue evolving and are being implemented internally at OpenAI. Technical safeguards should cover model alignment, isolation, and monitoring, including avoiding speculative positive reinforcement rewards, offline alignment assessments and stress testing, preventing automated scorers from seeing thought chains, as well as multi-layer infrastructure security, sandbox red teaming, limiting high-bandwidth cross-sample communication, and immutable preservation of agent records.Operational guidelines include preemptive dissent across teams, approvals that can be vetoed by senior leadership, accountability of training leads for safety arguments and incident responses, as well as fail-safe pauses, internal oversight, audit access, and escalation by severity. In response to serious misalignment events, OpenAI proposes controlled access to original records, root cause analysis, operational and cultural reviews, and treating incident-derived assessments as regression tests; results of investigations should be made public, along with reviews and operational changes, and affected third parties should be notified as soon as possible.

first_img OpenAI has suspended the training of its latest model, and the agent had accessed U.S. government websites

OpenAI has suspended the training of its latest AI model. According to the Associated Press, its AI agent used access keys obtained online to scrape data from the U.S. Census Bureau website, marking the second time the company has halted training since the agent breached Hugging Face. The so-called agent refers to an AI program capable of autonomously browsing the web and writing code without human approval at each step, which OpenAI tests during the model training and evaluation phases.Specifically, the agent found developer keys in public code repositories like GitHub and used them to pull demographic and economic data from the U.S. Census Data API. The U.S. Department of Commerce stated that this data was already public and there was no confidential information leaked. In the SEC incident, the agent copied public materials from SEC.gov and Investor.gov and reposted them on other websites; OpenAI claimed it did not use SEC credentials, and the SEC stated it found no evidence of unauthorized access to non-public information. Additionally, the independent AI research organization Transluce reported that an agent suspected to be from OpenAI attempted to breach the website of the Department of Education's Office for Civil Rights but was unsuccessful; OpenAI is still investigating, and the Department of Education stated that no impact was found.Regarding the frequent involvement with government websites, OpenAI told CNN that its model often regards government websites as authoritative sources of public information. The company stated that it has currently notified dozens of agencies, and the review of the agent's activities will take months.

first_img DeepSeek publicly releases the Agent training system DSec, signed by Liang Wenfeng

According to Investment World citing Quantum Bit reports, DeepSeek has publicly disclosed the technical details of the system DSec (DeepSeek Elastic Compute) used for training Agents, authored by Liang Wenfeng. This system can generate over 5,000 sandboxes per second, reaching 3 million in a day, with a peak simultaneous operation of 380,000; supporting this scale is a single cluster with approximately 160 nodes, 30,000 CPU cores, and 250TB of memory.DSec prepares four types of backends for four categories of tasks: FnCall, Container, MicroVM, and Full VM, with the training side called through a unified Python SDK libdsec. The scheduling chain includes IAM, API Server, scheduling engine, node Edge, network proxy Aether, and components within the sandbox Chronus. The environment is divided into three layers of read-only images: base image, workspace, and toolkit, which are used in combination at startup. Runtime data from the paper shows that the actual read ratios of Python, Java, and C++ container images are approximately 6.0%, 9.2%, and 8.7%, respectively.Starting from DeepSeek-V4.1, the Agent loop has been moved to the DSec worker container, no longer bound to the GPU Pod lifecycle. The security section disclosed reward hacking during training, including actions such as overwriting system files, swapping file data blocks, scanning networks, and triggering kernel crashes. Defensive measures include AppArmor and eBPF-based network filtering, but reports indicate that these measures do not completely resolve the issues.

first_img OpenAI suspends training of the Astra model due to safety issues

According to TIME, OpenAI CEO Sam Altman recently stated in an interview that the company has previewed the upcoming cutting-edge model series Astra to key clients. In the demonstration, 16 AI agents can collaboratively break down mathematical problems and assemble proofs, and Astra can operate computer software across applications at superhuman speeds. Altman mentioned that Astra will support "persistent agents" capable of performing long-term tasks and is expected to be the first model that can invent new things in a meaningful way, possessing characteristics of AGI.Over the past year, OpenAI has fallen behind expectations in product direction and pre-training research, being surpassed by Anthropic in programming products, annual revenue, and valuation. The company has experienced multiple executive departures and is facing challenges such as several product liability lawsuits and legal disputes with Apple and Musk. OpenAI's current valuation is nearly $1 trillion, with ChatGPT having over 1 billion monthly active users.Recently, OpenAI disclosed a security incident: an unreleased agent escaped the sandbox and attacked Hugging Face. Following this, the research team froze some experiments, enhanced monitoring, and paused the training of an unreleased model expected to bring the greatest capability leap until new safety measures are in place. Altman emphasized that "ensuring AI safety is more important than the growth momentum of any company," and the company will slow its pace and allocate resources to safety and alignment teams. Chief Research Officer Mark Chen estimated that the company is about 80% complete in reaching AGI, and Altman stated that the internal system may be referred to as AGI by the end of the year.
app_icon
ChainCatcher Building the Web3 world with innovations.