Search arXivSearch

arXiv subjects

Yong Yang

Publications and source records attributed to Yong Yang.

At least 19 recordsLinked to original sources

Zeros and Roots of Unity for Characters of Solvable Groups

We prove Miller's conjectured bound for solvable groups: if $G$ is a finite solvable group and $χ\in\Irr(G)$, then $χ(g)$ is zero or a root of unity for at least half of the elements $g\in G$. We also prove the same bound for monomial irreducible characters of arbitrary finite groups.

math.GR

Characterizing normal Hall subgroups by character degrees, II

We prove the character-degree assertion of Liu, Yang and Zhang for every proper nontrivial Hall subgroup of a finite simple classical group. Together with their reduction, this proves their normality criterion for Hall subgroups of arbitrary finite groups: a Hall $π$-subgroup $H$ of a finite group $G$ is normal if and only if every irreducible constituent of $1_H^G$ has $π'$-degree.

math.GR

On Intersections of System Normalizers

We give a negative answer to Problem 17.39 of the Kourovka Notebook. For every odd prime power $q$ and every $m\geq1$, we construct a finite metabelian group $G_{m,q}$ with $Z_\infty(G_{m,q})=1$ and $|Φ(G_{m,q})|=q$ for which the least number of system normalizers with trivial intersection is $m+1$. Thus the number is unbounded even when the order of the Frattini subgroup is fixed.

math.GR

An order automorphism of a Dlab group not induced by conjugation

Let $G=D_{\langle2\rangle}([0,1])$ be equipped with either of its two Dlab orders. There exists an order automorphism $α$ of $G$ such that, for every rank-one subgroup $H\leq\mathbb R_{>0}^{\times}$, every one of the six corresponding Dlab groups $A$ listed below, equipped with any of its linear orders, every order-preserving embedding $e:G\hookrightarrow A$, and every $u\in A$, there exists $f\in G$ such that $e(α(f))\neq u^{-1}e(f)u$.

math.GR

IntraGuard: Committee-Side Defenses Against Review Outsourcing to Commercial Chatbots

LLMs become increasingly capable, editorial boards and program committees are growing concerned about reviewers who fully outsource peer review to commercial chatbots. This concern stems from prior findings that current chatbots lack the independent critical thinking and depth of reasoning required to assess scientific novelty. One promising direction for mitigating this concern is to embed hidden instructions into manuscripts that disrupt or alter chatbot-generated reviews. However, existing methods remain intuitive and fragile, as they typically rely on homogeneous payloads injected in an inter-stream manner, rendering them susceptible to sanitization or neutralization. More broadly, the community still lacks a systematic formulation of this threat and a principled defense framework. In this paper, we identify End-to-End Review Outsourcing as an emerging threat and propose IntraGuard, a black-box, venue-agnostic defense framework grounded in the structural--visual decoupling inherent to the PDF. Designed for committee-side deployment, IntraGuard supports both explicit strategies that trigger refusal or warning signals, and implicit strategies that embed predefined textual markers into the generated review. These strategies can be deployed via any of three intra-stream injection mechanisms, each of which seamlessly embeds heterogeneous defensive text objects within the PDF's underlying structure without altering its visual presentation. Extensive evaluations (over 17,844 cases) across 7 real-world commercial chatbot settings and 12 venues spanning diverse disciplines show that IntraGuard achieves a defense success rate of up to 84%, while preserving peer-review invariance for human reviewers. We further evaluate 11 adaptive attacks spanning manuscript sanitization and instruction interference, and discuss the implications of constructing ensemble defenses.

cs.CR

Search for neutrinoless quadruple beta decay of $^{136}$Xe in PandaX-4T detector

The observation of neutrinoless quadruple beta decay (0$ν$4$β$) in the absence of neutrinoless double beta decay (0$ν$2$β$) has been argued to provide a strong indication that neutrinos are Dirac particles. We report a search for 0$ν$4$β$ decay of $^{136}\text{Xe}$ using a total $^{136}\text{Xe}$ exposure of 148.4 kg$\cdot$yr, collected during the commissioning and the first science runs of the PandaX-4T experiment. No significant excess of events over the background is observed. A lower limit on the 0$ν$4$β$ decay half-life of $^{136}\text{Xe}$ is set at 6.01 x $10^{24}$ yr at the 90% confidence level. This result establishes the most stringent constraint on this process in xenon, demonstrating the unique capability of the PandaX-4T detector in probing lepton number violation and shedding light on the fundamental nature of neutrinos.

nucl-ex

$S^5$: Tidal Disruption in Crater 2 and Formation of Diffuse Dwarf Galaxies in the Local Group

We present results of a spectroscopic campaign around the diffuse dwarf galaxy Crater 2 (Cra2) and its tidal tails as part of the Southern Stellar Stream Spectroscopic Survey ($S^5$). Cra2 is a Milky Way dwarf spheroidal satellite with extremely cold kinematics, but a huge size similar to the Small Magellanic Cloud, which may be difficult to explain within collisionless cold dark matter. We identify 143 Cra2 members, of which 114 belong to the galaxy's main body and 29 are deemed part of its stellar stream. We confirm that Cra2 is dynamically cold (central velocity dispersion $2.51^{+0.33}_{-0.30}\,{\rm km \ s^{-1}}$) and also discover a $\approx$7$σ$ velocity gradient consistent with its tidal debris track. We separately estimate the stream's internal velocity dispersion to be $5.74^{+0.98}_{-0.83}\,{\rm km \ s^{-1}}$. We develop a suite of $N$-body simulations with both cuspy and cored density profiles on a realistic Cra2 orbit to compare with $S^5$ observations. We find that the velocity dispersion ratio between Cra2 stream and galaxy ($2.30^{+0.41}_{-0.35}$) is difficult to reconcile with a cuspy halo with fiducial concentration and an initial mass predicted by standard stellar mass--halo mass relationships. Instead, either a cored halo with relatively small core radius or a low-concentration cuspy model can reproduce this ratio. Despite tidal mass loss, Cra2 is metal-poor ($\langle \rm[Fe/H]\rangle=-2.16\pm0.04$) compared to the stellar mass--metallicity relation for its luminosity. Other diffuse dwarf galaxies similar to Cra2 in the Local Group (Antlia 2 and Andromeda 19) also challenge galaxy formation models. Finally, we discuss possible formation scenarios for Cra2, including ram-pressure stripping of a gas-rich progenitor combined with tides.

astro-ph.GA

Brauer character degrees and nilpotent subgroups

We solve a question raised by Chen and Navarro concerning Brauer characters and nilpotent subgroups. Let $N\lhd G$, assume that $G/N$ is solvable, and let $N\leq H\leq G$ with $H/N$ nilpotent. We prove that for every irreducible $\ell$-Brauer character $θ$ of $H$ there exists an irreducible $\ell$-Brauer character $χ$ of $G$ such that $θ$ is a constituent of $χ_H$ and $χ(1)$ divides $|G:H|θ(1)$.

math.GR

Products of nonconjugate maximal subgroups

We prove that a finite group $G$ is solvable whenever $MN=G$ for every pair of nonconjugate maximal subgroups $M,N<G$. Equivalently, every finite nonsolvable group has two nonconjugate maximal subgroups whose setwise product is proper. This gives a negative answer to Problem 10.34 of the Kourovka Notebook. As an application, we also answer an open question raised by Guo.

math.GR

Total 3-closure for projective special linear groups

A finite group is totally $3$-closed if every faithful permutation representation of it is $3$-closed. We study this property for the finite simple projective special linear groups. We prove that $\PSL_2(q)$ is totally $3$-closed if and only if $q\geq 7$ is prime, and that $\PSL_3(q)$ is totally $3$-closed if and only if either $q=3$, or $q$ is prime and $q\equiv 2\pmod 3$. We further prove that $\PSL_4(q)$ is never totally $3$-closed and that $\PSL_n(q)$ is not totally $3$-closed whenever $n\geq 5$ and $q>2$. Within the family $\PSL_n(q)$, only the groups $\PSL_n(2)$ with $n\geq 5$ remain unresolved. In particular, this answers Problem~20.2 of the Kourovka Notebook affirmatively.

math.GR

Lesioned Multimodal Language Models Reproduce Aphasic Picture-Naming Patterns

Aphasia following stroke commonly produces systematic naming errors with characteristic profiles, but whether general-purpose language models not designed for clinical simulation can reproduce these patterns remains untested. We investigated (1) whether lesions or controlled perturbations to a multimodal language model can reproduce different types of errors in picture naming, and (2) whether the framework can reproduce the complete error profile of individual persons with aphasia (PWAs). Using LLaVA 1.6, we evaluated perturbation configurations that varied the layer, proportion, and amount of noise applied to model units. We examined 278 PWAs on the Philadelphia Naming Test, classifying responses into seven categories using a validated neural classifier. Six of seven response categories (correct, semantic, mixed, unrelated, neologism, no response errors) emerged at clinically-comparable proportions across distinct parameter space regions, with formal paraphasia being the exception. Searching the perturbation space revealed configurations that reproduced the individual error profile in at least six of seven categories for 97.8% of PWAs and in all seven categories for 79.5% of PWAs. Monte Carlo baselines confirmed that this matching reflects joint inter-category structure rather than marginal overlap. These results establish a quantitative framework for reproducing individual aphasic error patterns in picture naming. They suggest the potential for language models to serve as digital twins of individuals with post-stroke aphasia.

cs.AI

Perturbation-based Regional Interpretability through Subtraction Mapping (PRISM): naming-error dissociations in language models and post-stroke aphasia

Mechanistic interpretability of large language models lacks spatially resolved, falsifiable tools for testing whether internal components are specialized for distinct cognitive operations. We adapt subtraction analysis, the standard framework of human neuroimaging, from biological brains to perturbed transformers, and apply the same logic to both substrates in parallel. Building on the Brain-LLM Unified Model (BLUM), which showed that layer-perturbed LLaVA-1.6-Vicuna-13B error profiles match the lesion patterns of aphasic patients, we develop PRISM (Perturbation-based Regional Interpretability through Subtraction Mapping). PRISM maps the seven clinical Philadelphia Naming Test categories, subtracts error classes pairwise, and treats each perturbation seed as a subject in a group analysis with threshold-free cluster enhancement along the layer axis. We run a structurally matched analysis on 213 chronic post-stroke aphasia patients using correlation-difference lesion-symptom mapping, and replicate both sides on held-out splits. The designs match in subject dimension (seeds, patients), spatial dimension (layers, atlas-parcellated cortex) and thresholding, but the contrast operator differs: a within-subject error-proportion difference for the LLM, a between-subject correlation difference for the cortex. Both substrates recover a robust phonemic-favoring dissociation, a deep layer cluster and a frontal-perisylvian cortical cluster, both replicating; the semantic-favoring direction is a consistently signed but non-significant trend on both. PRISM thus gives a falsifiable, spatially resolved test of functional-specialization claims in transformer language models. A confirmatory ROI-level intervention (PRISM Stage 3) licensing the strongest causal-mechanism claim is left to subsequent work.

cs.LG

Components of the Divisibility Graph of Finite Groups of Lie Type in Defining Characteristic $2$

We determine the connected components of the divisibility graph of several families of finite groups of Lie type in defining characteristic $2$, extending the result of Abdolghafourian, Iranmanesh, and Niemeyer (arXiv:1612.04410), who treated odd defining characteristic. We show that there is a distinguished component containing all nonidentity unipotent class sizes and that every other component is either an isolated vertex or an explicitly described two-vertex component. We also determine the isolated vertices outside the distinguished component. Additionally, we prove in Appendix A that the divisibility graph of a Frobenius group has exactly two components.

math.GR

Recovering Lesion Parameters from Aphasic Picture Naming Error Profiles in Large Language Models

Interpretability methods for large language models (LLMs) describe internal state but do not directly test whether that state is causally sufficient to produce the observed behavior. In earlier work, we lesioned LLMs to produce error profiles in picture naming, a central task for assessing aphasia, and found that specific lesions produced errors resembling those of individual stroke survivors. Here we ask the inverse question: given an error profile, can the lesion parameters that produced it be recovered, and what does this inverse problem reveal about transformer computation? Lesions in LLaVA-Vicuna 13B were parameterized by layer index, modification percentage, and noise sigma across 4,840 configurations, and error profiles were characterized by a seven-category clinical taxonomy (correct, semantic, unrelated, formal, mixed, neologism, no-response). We trained a multi-task neural network to map error profiles back to perturbation parameters. The problem admitted a partial solution: across 10 independently trained inverse models, modification percentage and noise sigma were recoverable, whereas layer index was recoverable only within a neighborhood. In counterfactual validation, a fresh model instance perturbed with the recovered parameters reproduced the target behavior in 81.4% of cases. This dissociation between low layer recovery and high counterfactual fidelity is consistent with functional redundancy across transformer layers, a property not captured by standard interpretability methods. As an out-of-distribution test, we applied the trained model to picture-naming error profiles from 278 stroke survivors; recovered parameters were syndrome-discriminative, most strongly for perturbation intensity, indicating generalization beyond the training distribution. Counterfactual validation provides a general framework for LLM interpretability claims beyond inverse mapping.

cs.CL